anthropics / anthropics/claude-code

[Bug] Overly broad safeguard triggers false positive on legitimate internal IT administration and network diagnostics

オープン
#87,816 コメント 0 件 リアクション 0 件 担当者 0 名 GitHub で見る
area:model bug platform:macos stale
主要言語
Python
スター
145k
フォーク
23.1k
PR マージ指標
PR 指標を取得中

説明

**Bug Description**
Category: False positive — safeguard model-switch on legitimate work

I'm an internal IT administrator. I was diagnosing a slow camera feed at one of our own facilities, over our own private network, using our own authorized maintenance access — routine sysadmin work. I was also scoping on-prem occupancy analytics (counting how many people are in each area of our own facility) for staffing and equipment-ROI decisions.

Opus 5's safeguards flagged the thread and auto-switched me to Opus 4.8. The likely triggers were benign in context: "surveillance camera", "count people per area", read-only network inventory of our own subnets, and my reporting of a configuration weakness I found on our own equipment so we could fix it.

No offensive security, no third-party systems, no evasion — everything targeted assets my organization owns and I'm authorized to maintain. Please tune so authorized IT operations and on-prem camera analytics don't trip the model downgrade.

**Environment Info**
- Platform: darwin
- Terminal: iTerm.app
- Version: 2.1.235
- Feedback ID: dd9f5621-21e4-44cb-a715-9cc66dcd77a9

**Errors**
```json
[]
```

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

調査の方向性

The payload names no source files, tests, or entry points; begin by tracing the safeguard model-switch behavior associated with feedback ID dd9f5621-21e4-44cb-a715-9cc66dcd77a9. Reproduce the authorized IT diagnostics and on-prem camera-analytics context, then verify legitimate work no longer triggers the downgrade without weakening safeguards.

索引モデルが issue の本文から書いたものです。

評価

技術スタック
python
領域
security
issue の種類
バグ
難易度
5/5
見積もり時間
1週間以上
活発さ
活発
明瞭さ
説明が足りない
初心者へのやさしさ
25/100

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。