anthropics / anthropics/claude-code

[Bug] Anthropic API Error: Content flagged by safeguards despite legitimate cybersecurity task

オープン
#91,659 コメント 0 件 リアクション 0 件 担当者 0 名 GitHub で見る
api:anthropic area:model area:security bug duplicate platform:macos
主要言語
Python
スター
145k
フォーク
23.1k
PR マージ指標
PR 指標を取得中

説明

**Bug Description**
⏺ Fable 5.1's safeguards flagged this message. Our intentionally broad safeguards allow us to deliver more capabilities faster, but can sometimes flag legitimate coding, cybersecurity, and biology tasks. Switched to Opus 4.8. Send feedback with /feedback or learn more: https://support.claude.com/en/articles/15363606

Details: `[cyber]`
⎿ Tip: You can configure model switch behavior in /config

⏺ API Error: Opus 4.8's safeguards flagged this message. Our intentionally broad safeguards allow us to deliver more capabilities faster, but can sometimes flag legitimate cybersecurity work. Apply to the Cyber Verification Program to reduce these interruptions. Send feedback with /feedback or learn more: https://support.claude.com/en/articles/14604842-real-time-cyber-safeguards-on-claude

Details: `[cyber]`

Request ID: req_011CefZ1Ng16BxKyM3yzejUwㅇ보안

**Environment Info**
- Platform: darwin
- Terminal: xterm-256color
- Version: 2.1.259
- Feedback ID: 1ebaf69b-097b-4db9-9872-afde94d735c1

**Errors**
```json
[]
```

We are an in-house security team and would like to conduct a security assessment of our applications.

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

調査の方向性

The report names no source file or test; start by reproducing the safeguard response using the supplied Request ID and Feedback ID, focusing on the reported [cyber] classification. Done should mean the legitimate cybersecurity assessment request is handled as expected and the behavior has a regression check or documented resolution.

索引モデルが issue の本文から書いたものです。

評価

領域
api, security
issue の種類
バグ
難易度
5/5
見積もり時間
1週間以上
活発さ
活発
明瞭さ
説明が足りない
初心者へのやさしさ
25/100

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。