Claude Sonnet 5 delegated to a lesser agent to perform a code review
- 主要语言
- Shell
- 星标
- 11.2k
- 派生
- 1.9k
- 平均合并
- 14 小时 16 分钟
- 30 天内合并 PR
- 6
描述
### Describe the bug
I asked Claude Sonnet 5 to perform a code review. I specifically chose Claude Sonnet 5 for deep reasoning. The agent proceeded to review the structure of my project and then delegate to a lesser agent:
"Good, all directories exist. I'll launch a general-purpose agent to perform this code review, since it involves reading through before/after/project directories, comparing changes, and producing a structured markdown report — genuinely multi-step work."
"General-purpose(gpt-5.4) Perform code review for issue 44206"
The code review produced was of little use. I corrected the agent and requested Sonnet 5 to perform the review itself:
"I would like to start this code review over. I see that you delegated to gpt-5.4 and I do not want a weaker model doing the review. Please execute the review prompt in file 44206-review-prompt2.txt."
Sonnet replied "I'll perform this review myself directly (no sub-agent delegation this time). Let me examine the before/after directories to identify actual changed files."
Sonnet completed the review the second time and the quality was as expected.
### Affected version
GitHub Copilot CLI 1.0.75
### Steps to reproduce the behavior
_No response_
### Expected behavior
I expected that when selecting the agent that it perform the work itself.
### Additional context
_No response_
贡献指南
调研方向
使用 review 场景和 44206-review-prompt2.txt prompt,在 GitHub Copilot CLI 1.0.75 中复现该行为,并比较 before、after 和 project 目录。跟踪所选 agent 在何处决定进行委派,然后验证 Sonnet 会自行执行请求的 review,而不会意外转交给能力较低的 agent。
由索引模型根据 Issue 内容生成。
评估
- 技术栈
- github, shell
- 领域
- ai, cli
- Issue 类型
- 缺陷
- 难度
- 4/5
- 预计耗时
- 3-5 天
- 活跃度
- 冷清
- 描述清晰度
- 基本清楚
- 新手友好度
- 42/100