Request consumption appears abnormally high — possible double/triple counting
- Lenguaje dominante
- Shell
- Estrellas
- 11.2k
- Forks
- 1.9k
- Merge medio
- 14 h 16 min
- PR fusionados (30 d)
- 6
Descripción
### Describe the bug
I'm using the 1x multiplier model (GPT-5.4), but request consumption is behaving
as if a 3x multiplier model is selected. Each interaction burns through quota at
roughly 3x the expected rate, despite explicitly choosing the lower-cost model tier.
This started approximately 3 days ago with no changes on my end, same model
selection, same workflows, same shell environment.
### Affected version
GitHub Copilot CLI 1.0.22.
### Steps to reproduce the behavior
Set model to GPT-5.4
Run a typical planing - implementing workflow, says costs 1x req, but feels like 3x, %'s are increasing way too fast.
### Expected behavior
Expected: 1 request deducted per interaction (1x model)
Actual: ~3 requests deducted per interaction — consumption matches a 3x model
### Additional context
_No response_
Guía de contribución
Línea de trabajo
No files or tests are named. Start by reproducing the reported workflow in GitHub Copilot CLI 1.0.22 with GPT-5.4, then trace how model selection and request consumption are reported. Done means confirming whether one interaction deducts about three requests and identifying the cause or documenting why the quota differs from the expected 1x rate.
Escrito por el modelo de indexación a partir del texto del issue.
Evaluación
- Stack tecnológico
- shell
- Área
- cli
- Tipo de issue
- Error
- Dificultad
- 3/5
- Tiempo estimado
- 1-2 días
- Estado de actividad
- Tranquilo
- Claridad
- Bastante claro
- Aptitud para principiantes
- 48/100