Opus 4.7 small context window triggers auto-compact too frequently
- 主要語言
- Shell
- 星號
- 11.2k
- 分支
- 1.9k
- 平均合併
- 14 小時 16 分鐘
- 30 天內合併 PR
- 6
描述
### Describe the bug
When using `Opus 4.7` with a Copilot Pro+ subscription, the effective available context window appears much smaller than comparable models like `GPT 5.4` under the same conditions.
In practice, this causes auto-compact to trigger very frequently, including multiple times within a single prompt/session, which makes the model difficult to use even for medium-complexity tasks.
This is especially noticeable because Opus-class models are generally better suited for more complex tasks, which often require larger working context. I understand that the 1M tokens context allocation may not be feasible for cost or product reasons, but the current limit seems too restrictive to be practical.
In my case, a large portion of the context appears to be consumed by `System/Tools`, leaving much less room for actual task context than with `GPT 5.4`.
`Opus 4.7` context
`GPT 5.4` context (for comparison using the same plugins and tools on the same machine)
### Affected version
GitHub Copilot CLI 1.0.36
### Steps to reproduce the behavior
1. Start a fresh Copilot CLI session on a Pro+ subscription.
2. Switch to `Opus 4.7` model.
3. Run a simple prompt to populate the context and observe that a large share of the context is already occupied by `System/Tools`.
4. In the same setup, read a moderate amount of content (for example, around 10 markdown files of roughly 200-300 lines / 15-20 KB each) and ask for a summary (or any other medium effort task).
5. Observe that the context fills up and auto-compact triggers quickly, often multiple times during the same session or even during a single prompt.
6. Repeat the same workflow with `GPT 5.4` (Or any other comparable model like Sonnet) and compare effective remaining context and auto-compact frequency.
### Expected behavior
1. `Opus 4.7` should reserve a smaller portion of the window for `System/Tools`, closer to `GPT 5.4` (Or any other comparable model like Sonnet) under the same setup.
2. A medium-complexity prompt should fit without repeated auto-compact during normal use.
3. If the model must have a smaller context budget than GPT 5.4, it should still be large enough to handle moderate multi-file workflows reliably.
### Additional context
- OS: Windows 11
- Shell: PowerShell
- Subscription: Copilot Pro+
- Comparison was made using the same plugins/tools setup as `GPT 5.4`
貢獻指南
研究方向
No source file, test, or entry point is named. Start by reproducing the comparison in Copilot CLI 1.0.36 with Opus 4.7 and GPT 5.4 using the same plugins and tools, then trace how System/Tools context usage and auto-compact thresholds are selected. Done means medium multi-file prompts no longer trigger repeated auto-compaction, or the model-specific limitation is clearly handled.
由索引模型根據 Issue 內容生成。
評估
- 技術堆疊
- shell
- 領域
- ai, cli
- Issue 類型
- 缺陷
- 難度
- 4/5
- 預估耗時
- 3-5 天
- 活躍度
- 冷清
- 描述清晰度
- 需要釐清
- 新手友好度
- 35/100