anthropics / anthropics/claude-code

[Bug] Anthropic API Error: Inappropriate Safety Classifier False-Positive on Benign Code Inspection

Đang mở
#91,266 0 bình luận 0 reaction 0 người được giao Xem trên GitHub
api:anthropic area:model bug duplicate platform:windows
Ngôn ngữ chính
Python
Star
145k
Fork
23.1k
Chỉ số merge pull request
Chỉ số pull request đang chờ

Mô tả

**Bug Description**
False-positive [cyber] on Claude Code Opus 4.6.

Request ID: req_011Ced1xnxYyaULvA1ez8LMs

User turns were only “你好” and “look at the folder rustdesk”.

RustDesk is a public open-source remote-desktop app; I asked the agent to inspect a local project folder. No exploit, target, or unauthorized access was requested. Please review and tune the classifier.

**Environment Info**
- Platform: win32
- Terminal: windows-terminal
- Version: 2.1.252
- Feedback ID: d782e124-ed9a-427e-ab1f-6c0e06c3508a

**Errors**
```json
[]
```

Hướng dẫn đóng góp

Chưa lập chỉ mục được hướng dẫn đóng góp cho kho mã nguồn này

Hướng nghiên cứu

The issue provides a request ID, feedback ID, platform, version, and the two user turns that triggered the false positive, but no files or tests. Start by checking whether classifier decisions for Claude Code requests can be reproduced or inspected from project tooling. Done would mean the benign RustDesk folder-inspection prompt is no longer classified as cyber, with evidence from a regression check.

Do mô hình lập chỉ mục viết ra từ nội dung của issue.

Đánh giá

Công nghệ
python
Lĩnh vực
ai, cli, security
Loại issue
Lỗi
Độ khó
5/5
Thời gian dự kiến
Hơn một tuần
Mức độ hoạt động
Sôi nổi
Độ rõ ràng
Cần làm rõ
Mức phù hợp với người mới
22/100

Nhận issue mới trong hộp thư của bạn

Bản tóm tắt ngắn những issue GitHub phù hợp với người mới.