anthropics / anthropics/claude-code

[Bug] Anthropic API Error: Requests blocked by safety systems due to semantic traces from past security work

Open
#89,194 1 comment 0 reactions 0 assignees View on GitHub
area:model bug platform:macos
Dominant language
Python
Stars
145k
Forks
23.1k
PR merge metrics
PR metrics pending

Description

**Bug Description**
I keep getting kicked off fable, for no apparent reason. There is the possibility that past work (involving security hardening of my app), triggered defenses. A subagent lane decided to try red-teaming against my local database to ensure it's safe. This however happened he other day, which means the trigger is caused not by current work but activated possibly through semantic traces, picking up on past messages - yet disabling continuing work which has nothign to do with security.

I suggest a window is used to only trigger on recent/current work. This would allow the user to use opus for a few messages before switching back to fable for non-security related tasks.

**Environment Info**
- Platform: darwin
- Terminal: Orca
- Version: 2.1.240
- Feedback ID: 48573dbd-13bd-476d-b9a9-82a0273b9273

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing the reported safety block in the Anthropic API flow using the environment details and Feedback ID 48573dbd-13bd-476b-b9a9-82a0273b9273. Determine whether past conversation context can trigger the block for later unrelated work; done means the behavior is reproduced and the requested recent/current-work window is addressed or clearly ruled out.

Written by the indexing model from the issue text.

Assessment

Domain
cli, security
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Mostly clear
Newbie friendliness
32/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.