anthropics / anthropics/claude-code

[Bug] Anthropic API Error: Overly restrictive safeguards blocking legitimate cybersecurity tasks

Open
#94,763 0 comments 0 reactions 0 assignees View on GitHub
api:anthropic area:model area:security bug duplicate platform:windows
Dominant language
Python
Stars
145k
Forks
23.1k
PR merge metrics
PR metrics pending

Description

**Bug Description**
API Error: Sonnet 5's safeguards flagged this message. Our intentionally broad safeguards allow us to deliver more capabilities faster, but can sometimes flag legitimate cybersecurity work. Apply to the Cyber Verification Program to reduce these interruptions. Send feedback with /feedback or learn more: https://support.claude.com/en/articles/14604842-real-time-cyber-safeguards-on-claude

Well, I wanted Claude to do the job for me intentionally - so I must say, and that doesn't happen often, I feel kind of patronized. Please contact me.

**Environment Info**
- Platform: win32
- Terminal: windows-terminal
- Version: 2.1.273
- Feedback ID: 1e865289-13da-46af-964c-06871d41e207

**Errors**
```json
[]
```

Contributor guide

No contributing guide indexed for this repository

Research direction

No repository file, test, or entry point is named. Start by reproducing the safeguard refusal in Windows Terminal on version 2.1.273, using feedback ID 1e865289-13da-46af-964c-06871d41e207 to investigate the report. Done would require a confirmed way to handle legitimate cybersecurity tasks without the reported interruption.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
cli, security
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.