anthropics / anthropics/claude-code

Fable 5 safeguard false-positives derail long agent sessions

Open
#87,718 0 comments 0 reactions 0 assignees View on GitHub
area:model bug stale
Dominant language
Python
Stars
145k
Forks
23.1k
PR merge metrics
PR metrics pending

Description

During a legitimate dev session (applying a database migration to my own product's prod DB), Fable 5's safeguards flagged routine messages twice (02:59Z and 03:08Z on 2026-07-27). The first flag silently switched my session to Opus 4.8 (despite switchModelsOnFlag: false); the second erred every retry, and because the flagged content remained in context the session could not recover — I lost roughly an hour and two work items.

Separately: in the mobile app's live voice mode, tool calls play a loud repeating progress chime with no mute setting, which makes long-running tool loops unusable by ear.

Requests:
1. Safeguard flags should not fire on credential-indirection dev-ops text, or should degrade more gracefully than silent model switching.
2. A visible "this turn was flagged — rephrase it" recovery hint.
3. A sounds toggle for live-mode tool calls.

Contributor guide

No contributing guide indexed for this repository

Research direction

No files, tests, or entry points are named. Separate the safeguard behavior, flagged-turn recovery, and mobile live-mode audio requests, then locate the relevant subsystems before checking how each behavior is currently handled. Done means the reported false-positive flow recovers without unwanted model switching or repeated failure, and live tool-call sounds can be muted.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
cli, mobile-dev, security
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Mostly clear
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.