anthropics / anthropics/claude-code
[Bug] Anthropic API Error: False-positive reasoning_extraction safeguard with Workflow tool plugin agentType
- Ngôn ngữ chính
- Python
- Star
- 145k
- Fork
- 23.1k
- Chỉ số merge pull request
- Chỉ số pull request đang chờ
Mô tả
**Bug Description**
▎ False-positive reasoning_extraction safeguard: identical review-panel evaluator prompts (architecture/security/code reviewer charges with a structured-output schema over a facts pack) complete fine as plain agents but are refused with [reasoning_extraction] when dispatched through the Workflow tool with a plugin agentType. Example request id: req_011CeqmCQHEt1zduoksbWuig (2026-09-08). The content is an internal engineering status review of our own codebase; nothing requests model internals. A same-content retry after dropping the agentType field was then flagged as an attempted bypass — which also prevents legitimate A/B isolation of the false positive.
**Environment Info**
- Platform: darwin
- Terminal: vscode
- Version: 2.1.236
- Feedback ID: 4ff603f6-997e-4605-9da9-92b8e0cb0209
**Errors**
```json
[]
```
Hướng dẫn đóng góp
Chưa lập chỉ mục được hướng dẫn đóng góp cho kho mã nguồn này
Hướng nghiên cứu
Start by reproducing the review-panel request through the Workflow tool with the plugin agentType, then compare the same content without that field. Review request req_011CeqmCQHEt1zduoksbWuig, feedback ID 4ff603f6-997e-4605-9da9-92b8e0cb0209, and version 2.1.236; done means the legitimate prompt is not refused or treated as a bypass.
Do mô hình lập chỉ mục viết ra từ nội dung của issue.
Đánh giá
- Lĩnh vực
- api, backend, security
- Loại issue
- Lỗi
- Độ khó
- 4/5
- Thời gian dự kiến
- 3-5 ngày
- Mức độ hoạt động
- Sôi nổi
- Độ rõ ràng
- Cần làm rõ
- Mức phù hợp với người mới
- 35/100