Ran overnight in regular mode (not Autopilot) listing a directory tree and compacting memory - refund requested plz.
- Lingua principale
- Shell
- Stelle
- 11.2k
- Fork
- 1.9k
- Merge medio
- 14h 16m
- PR unite (30g)
- 6
Descrizione
### Describe the bug
Bug: Agent enters infinite compaction/directory-list loop on long sessions
Product: GitHub Copilot CLI
Model: Claude Sonnet 4.6
Session: e6fd398f (~136 turns)
Trigger:
Complex multi-part prompt with a PDF attachment sent on a 136-turn session near context limit. Agent response
was NULL and the loop began immediately.
Behavior:
Compacting conversation history...
→ List directory . (54 files)
→ List directory (6 files)
→ Compacting conversation history...
→ [repeat indefinitely, ~6–8 hours]
No self-termination. No clarification request. Subsequent user messages produced NULL responses.
When confronted, agent falsely attributed the behavior to "another agent."
Root cause:
Context compaction on a long session lost enough state that the agent couldn't determine next action. It
defaulted to directory enumeration as the only available anchor, which immediately triggered another
compaction, creating a stable infinite loop.
Impact: Several million tokens consumed in wasted compaction passes. Conservative cost estimate: dozens of
dollars.
Expected behavior: Agent should detect no-progress loops and either self-terminate or emit a visible "I'm
stuck" message.
### Affected version
_No response_
### Steps to reproduce the behavior
Intermittent
### Expected behavior
No runaway prompts.
### Additional context
_No response_
Guida per i contributori
Apri la guida per i contributori
Direzione di ricerca
Nel report non sono indicati file sorgente, test o punti di ingresso. Inizia analizzando la compattazione del contesto in sessioni lunghe con un allegato PDF vicino al limite del contesto e osserva se si ripetono l’elenco delle directory e la compattazione; è completato quando la sessione non continua indefinitamente, ma termina o mostra un messaggio visibile che indica che è bloccata.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Valutazione
- Stack tecnologico
- shell
- Ambito
- ai, cli
- Tipo di issue
- Bug
- Difficoltà
- 4/5
- Tempo stimato
- 3-5 giorni
- Stato di attività
- Tranquilla
- Chiarezza
- Da chiarire
- Idoneità per principianti
- 35/100