Ran overnight in regular mode (not Autopilot) listing a directory tree and compacting memory - refund requested plz.
- Langage dominant
- Shell
- Étoiles
- 11.2k
- Forks
- 1.9k
- Merge moyen
- 14 h 16 min
- PR mergées (30 j)
- 6
Description
### Describe the bug
Bug: Agent enters infinite compaction/directory-list loop on long sessions
Product: GitHub Copilot CLI
Model: Claude Sonnet 4.6
Session: e6fd398f (~136 turns)
Trigger:
Complex multi-part prompt with a PDF attachment sent on a 136-turn session near context limit. Agent response
was NULL and the loop began immediately.
Behavior:
Compacting conversation history...
→ List directory . (54 files)
→ List directory (6 files)
→ Compacting conversation history...
→ [repeat indefinitely, ~6–8 hours]
No self-termination. No clarification request. Subsequent user messages produced NULL responses.
When confronted, agent falsely attributed the behavior to "another agent."
Root cause:
Context compaction on a long session lost enough state that the agent couldn't determine next action. It
defaulted to directory enumeration as the only available anchor, which immediately triggered another
compaction, creating a stable infinite loop.
Impact: Several million tokens consumed in wasted compaction passes. Conservative cost estimate: dozens of
dollars.
Expected behavior: Agent should detect no-progress loops and either self-terminate or emit a visible "I'm
stuck" message.
### Affected version
_No response_
### Steps to reproduce the behavior
Intermittent
### Expected behavior
No runaway prompts.
### Additional context
_No response_
Guide de contribution
Ouvrir le guide de contribution
Piste de recherche
Le rapport ne mentionne aucun fichier source, test ou point d’entrée. Commencez par étudier la compaction du contexte lors d’une session longue avec une pièce jointe PDF proche de la limite de contexte, et observez si des listings de répertoires et une compaction se répètent ; c’est terminé lorsque la session ne s’exécute pas indéfiniment, mais se termine ou affiche un message visible indiquant qu’elle est bloquée.
Rédigé par le modèle d'indexation à partir du texte de l'issue.
Évaluation
- Stack technique
- shell
- Domaine
- ai, cli
- Type d'issue
- Bug
- Difficulté
- 4/5
- Temps estimé
- 3-5 jours
- Activité
- Calme
- Clarté
- À clarifier
- Accessibilité débutants
- 35/100