anthropics / anthropics/claude-code

[Bug] Anthropic API Error: Opus 5 safeguards blocking legitimate cybersecurity tasks

Closed
#94,401 0 comments 0 reactions 0 assignees View on GitHub
area:model bug duplicate platform:macos
Dominant language
Python
Stars
145k
Forks
23.1k
PR merge metrics
PR metrics pending

Description

**Bug Description**
API Error: Opus 5's safeguards flagged this message (https://www.anthropic.com/legal/aup). Our intentionally broad safeguards allow us to deliver more capabilities faster, but can sometimes flag legitimate coding, cybersecurity, and biology tasks. Claude Code can't respond to this message with Opus 5.

Double press esc to edit your last message, or try a different model with /model.

Send feedback with /feedback or learn more: https://support.claude.com/en/articles/16049681

Details: `[cyber]`

Context: [9]: https://academy.hackthebox.com/preview/certifications?utm_source=chatgpt.com "Cybersecurity Certifications | Prove Practical Skills. Get Hired."

**Environment Info**
- Platform: darwin
- Terminal: ghostty

😆

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reproducing the reported Opus 5 safeguard error in Claude Code for the cybersecurity task described, then compare it with a different model using /model. The report does not include the blocked prompt, source file, or test; done would require confirming that legitimate cybersecurity work is no longer incorrectly blocked.

Written by the indexing model from the issue text.

Assessment

Domain
cli, security
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.