Claude Sonnet 5 delegated to a lesser agent to perform a code review
まだ誰も着手していません。
- 主要言語
- Shell
- スター
- 11.2k
- フォーク
- 1.9k
- 平均マージ
- 14時間 16分
- マージ済み PR(30日)
- 6
説明
Describe the bug
I asked Claude Sonnet 5 to perform a code review. I specifically chose Claude Sonnet 5 for deep reasoning. The agent proceeded to review the structure of my project and then delegate to a lesser agent:
"Good, all directories exist. I'll launch a general-purpose agent to perform this code review, since it involves reading through before/after/project directories, comparing changes, and producing a structured markdown report — genuinely multi-step work."
"General-purpose(gpt-5.4) Perform code review for issue 44206"
The code review produced was of little use. I corrected the agent and requested Sonnet 5 to perform the review itself:
"I would like to start this code review over. I see that you delegated to gpt-5.4 and I do not want a weaker model doing the review. Please execute the review prompt in file 44206-review-prompt2.txt."
Sonnet replied "I'll perform this review myself directly (no sub-agent delegation this time). Let me examine the before/after directories to identify actual changed files."
Sonnet completed the review the second time and the quality was as expected.
Affected version
GitHub Copilot CLI 1.0.75
Steps to reproduce the behavior
No response
Expected behavior
I expected that when selecting the agent that it perform the work itself.
Additional context
No response
コントリビューションガイド
はじめの一歩
- issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
- 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
- リポジトリをフォークし、ブランチを切って変更します。
- issue 番号を参照したプルリクエストを送ります。
調査の方向性
GitHub Copilot CLI 1.0.75 で、レビューシナリオと 44206-review-prompt2.txt プロンプトを使い、before、after、project ディレクトリを比較して動作を再現する。選択されたエージェントがどこで委譲を判断するのかを追跡し、その後、Sonnet が意図しない下位エージェントへの引き渡しを行わずに、要求されたレビューを自ら実行することを確認する。
索引モデルが issue の本文から書いたものです。
評価
- 技術スタック
- github, shell
- 領域
- ai, cli
- issue の種類
- バグ
- 難易度
- 4/5
- 見積もり時間
- 3〜5日
- 活発さ
- 静か
- 明瞭さ
- おおむね明確
- 初心者へのやさしさ
- 42/100