Request consumption appears abnormally high — possible double/triple counting
- Dominant language
- Shell
- Stars
- 11.2k
- Forks
- 1.9k
- Avg merge
- 14h 16m
- Merged PRs (30d)
- 6
Description
### Describe the bug
I'm using the 1x multiplier model (GPT-5.4), but request consumption is behaving
as if a 3x multiplier model is selected. Each interaction burns through quota at
roughly 3x the expected rate, despite explicitly choosing the lower-cost model tier.
This started approximately 3 days ago with no changes on my end, same model
selection, same workflows, same shell environment.
### Affected version
GitHub Copilot CLI 1.0.22.
### Steps to reproduce the behavior
Set model to GPT-5.4
Run a typical planing - implementing workflow, says costs 1x req, but feels like 3x, %'s are increasing way too fast.
### Expected behavior
Expected: 1 request deducted per interaction (1x model)
Actual: ~3 requests deducted per interaction — consumption matches a 3x model
### Additional context
_No response_
Contributor guide
Research direction
No files or tests are named. Start by reproducing the reported workflow in GitHub Copilot CLI 1.0.22 with GPT-5.4, then trace how model selection and request consumption are reported. Done means confirming whether one interaction deducts about three requests and identifying the cause or documenting why the quota differs from the expected 1x rate.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- shell
- Domain
- cli
- Issue type
- Bug
- Difficulty
- 3/5
- Estimated time
- 1-2 days
- Activity status
- Quiet
- Clarity
- Mostly clear
- Newbie friendliness
- 48/100