github / github/copilot-cli

Request consumption appears abnormally high — possible double/triple counting

Open
#2,626 2 comments 0 reactions 0 assignees View on GitHub
area:models
Dominant language
Shell
Stars
11.2k
Forks
1.9k
Avg merge
14h 16m
Merged PRs (30d)
6

Description

### Describe the bug

I'm using the 1x multiplier model (GPT-5.4), but request consumption is behaving
as if a 3x multiplier model is selected. Each interaction burns through quota at
roughly 3x the expected rate, despite explicitly choosing the lower-cost model tier.

This started approximately 3 days ago with no changes on my end, same model
selection, same workflows, same shell environment.

### Affected version

GitHub Copilot CLI 1.0.22.

### Steps to reproduce the behavior

Set model to GPT-5.4
Run a typical planing - implementing workflow, says costs 1x req, but feels like 3x, %'s are increasing way too fast.

### Expected behavior

Expected: 1 request deducted per interaction (1x model)
Actual: ~3 requests deducted per interaction — consumption matches a 3x model

### Additional context

_No response_

Contributor guide

Open the contributing guide

Research direction

No files or tests are named. Start by reproducing the reported workflow in GitHub Copilot CLI 1.0.22 with GPT-5.4, then trace how model selection and request consumption are reported. Done means confirming whether one interaction deducts about three requests and identifying the cause or documenting why the quota differs from the expected 1x rate.

Written by the indexing model from the issue text.

Assessment

Tech stack
shell
Domain
cli
Issue type
Bug
Difficulty
3/5
Estimated time
1-2 days
Activity status
Quiet
Clarity
Mostly clear
Newbie friendliness
48/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.