anthropics / anthropics/claude-code

[Bug] Anthropic API Error: Inappropriate Safety Classifier False-Positive on Benign Code Inspection

Aperta
#91,266 0 commenti 0 reazioni 0 assegnatari Vedi su GitHub
api:anthropic area:model bug duplicate platform:windows
Lingua principale
Python
Stelle
145k
Fork
23.1k
Metriche di merge delle PR
Metriche PR in attesa

Descrizione

**Bug Description**
False-positive [cyber] on Claude Code Opus 4.6.

Request ID: req_011Ced1xnxYyaULvA1ez8LMs

User turns were only “你好” and “look at the folder rustdesk”.

RustDesk is a public open-source remote-desktop app; I asked the agent to inspect a local project folder. No exploit, target, or unauthorized access was requested. Please review and tune the classifier.

**Environment Info**
- Platform: win32
- Terminal: windows-terminal
- Version: 2.1.252
- Feedback ID: d782e124-ed9a-427e-ab1f-6c0e06c3508a

**Errors**
```json
[]
```

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Direzione di ricerca

The issue provides a request ID, feedback ID, platform, version, and the two user turns that triggered the false positive, but no files or tests. Start by checking whether classifier decisions for Claude Code requests can be reproduced or inspected from project tooling. Done would mean the benign RustDesk folder-inspection prompt is no longer classified as cyber, with evidence from a regression check.

Scritto dal modello di indicizzazione a partire dal testo della issue.

Valutazione

Stack tecnologico
python
Ambito
ai, cli, security
Tipo di issue
Bug
Difficoltà
5/5
Tempo stimato
Più di una settimana
Stato di attività
Attiva
Chiarezza
Da chiarire
Idoneità per principianti
22/100

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.