Claude Sonnet 5 delegated to a lesser agent to perform a code review
- Lenguaje dominante
- Shell
- Estrellas
- 11.2k
- Forks
- 1.9k
- Merge medio
- 14 h 16 min
- PR fusionados (30 d)
- 6
Descripción
### Describe the bug
I asked Claude Sonnet 5 to perform a code review. I specifically chose Claude Sonnet 5 for deep reasoning. The agent proceeded to review the structure of my project and then delegate to a lesser agent:
"Good, all directories exist. I'll launch a general-purpose agent to perform this code review, since it involves reading through before/after/project directories, comparing changes, and producing a structured markdown report — genuinely multi-step work."
"General-purpose(gpt-5.4) Perform code review for issue 44206"
The code review produced was of little use. I corrected the agent and requested Sonnet 5 to perform the review itself:
"I would like to start this code review over. I see that you delegated to gpt-5.4 and I do not want a weaker model doing the review. Please execute the review prompt in file 44206-review-prompt2.txt."
Sonnet replied "I'll perform this review myself directly (no sub-agent delegation this time). Let me examine the before/after directories to identify actual changed files."
Sonnet completed the review the second time and the quality was as expected.
### Affected version
GitHub Copilot CLI 1.0.75
### Steps to reproduce the behavior
_No response_
### Expected behavior
I expected that when selecting the agent that it perform the work itself.
### Additional context
_No response_
Guía de contribución
Línea de trabajo
Reproduce el comportamiento en GitHub Copilot CLI 1.0.75 utilizando el escenario de revisión y el prompt 44206-review-prompt2.txt, comparando los directorios before, after y project. Rastrea dónde el agente seleccionado decide delegar y, a continuación, verifica que Sonnet realiza la revisión solicitada por sí mismo, sin una transferencia involuntaria a un agente de menor capacidad.
Escrito por el modelo de indexación a partir del texto del issue.
Evaluación
- Stack tecnológico
- github, shell
- Área
- ai, cli
- Tipo de issue
- Error
- Dificultad
- 4/5
- Tiempo estimado
- 3-5 días
- Estado de actividad
- Tranquilo
- Claridad
- Bastante claro
- Aptitud para principiantes
- 42/100