Claude Sonnet 5 delegated to a lesser agent to perform a code review
Dieses Issue hat noch niemand übernommen.
- Vorherrschende Sprache
- Shell
- Sterne
- 11.2k
- Forks
- 1.9k
- Ø Merge
- 14 Std. 16 Min.
- Gemergte PRs (30 T.)
- 6
Beschreibung
Describe the bug
I asked Claude Sonnet 5 to perform a code review. I specifically chose Claude Sonnet 5 for deep reasoning. The agent proceeded to review the structure of my project and then delegate to a lesser agent:
"Good, all directories exist. I'll launch a general-purpose agent to perform this code review, since it involves reading through before/after/project directories, comparing changes, and producing a structured markdown report — genuinely multi-step work."
"General-purpose(gpt-5.4) Perform code review for issue 44206"
The code review produced was of little use. I corrected the agent and requested Sonnet 5 to perform the review itself:
"I would like to start this code review over. I see that you delegated to gpt-5.4 and I do not want a weaker model doing the review. Please execute the review prompt in file 44206-review-prompt2.txt."
Sonnet replied "I'll perform this review myself directly (no sub-agent delegation this time). Let me examine the before/after directories to identify actual changed files."
Sonnet completed the review the second time and the quality was as expected.
Affected version
GitHub Copilot CLI 1.0.75
Steps to reproduce the behavior
No response
Expected behavior
I expected that when selecting the agent that it perform the work itself.
Additional context
No response
Beitragsleitfaden
Erste Schritte
- Lies das ganze Issue und danach den Beitragsleitfaden des Projekts.
- Schreib ins Issue, dass du es übernimmst — das erspart doppelte Arbeit.
- Forke das Repository und arbeite in einem Branch.
- Öffne einen Pull Request, der die Issue-Nummer nennt.
Rechercherichtung
Reproduziere das Verhalten in GitHub Copilot CLI 1.0.75 anhand des Review-Szenarios und des Prompts 44206-review-prompt2.txt und vergleiche dabei die Verzeichnisse before, after und project. Verfolge, wo der ausgewählte Agent entscheidet, zu delegieren, und prüfe anschließend, dass Sonnet das angeforderte Review selbst durchführt, ohne eine unbeabsichtigte Übergabe an einen weniger leistungsfähigen Agenten.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Bewertung
- Tech-Stack
- github, shell
- Bereich
- ai, cli
- Issue-Typ
- Bug
- Schwierigkeit
- 4/5
- Geschätzter Aufwand
- 3-5 Tage
- Aktivitätsstatus
- Ruhig
- Klarheit
- Größtenteils klar
- Anfängerfreundlichkeit
- 42/100