GOAT Plan: DeepSeek V4 Flash Cache Read billed ~53% higher than published rate (off-peak usage)
Dieses Issue hat noch niemand übernommen.
Bewertung
- Schwierigkeit
- 4/5
- Geschätzter Aufwand
- 3-5 Tage
- Anfängerfreundlichkeit
- 35/100
Rechercherichtung
Beginne mit der GOAT-Preistabelle und der Nutzungsaufschlüsselung im Dashboard für DeepSeek V4 Flash und untersuche anschließend, ob die vollständige Trace-Historie über eine API oder einen Log-Export verfügbar ist. Vergleiche die Abrechnung für Cache Read während des gemeldeten Nebenzeitraums mit dem veröffentlichten Tarif; abgeschlossen ist die Untersuchung, wenn die Abweichung korrigiert oder ihre Abrechnungsgrundlage eindeutig erklärt ist.
Vom Indexierungsmodell aus dem Issue-Text verfasst.
Beschreibung
Summary
Description
I've identified a billing discrepancy on the GOAT plan when using DeepSeek V4 Flash. The actual cost for Cache Read tokens is significantly higher than the published rate in the official pricing table, while Input and Output costs match the table exactly.
Expected Behavior
Based on the official GOAT pricing table for DeepSeek V4 Flash:
- Input: $0.22/M
- Output: $0.66/M
- Cache Read: $0.007/M
With my usage (all during off-peak hours):
-
Input: 2.5M × $0.22 = $0.55
-
Output: 0.7874M × $0.66 = $0.5197
-
Cache Read: 83.1M × $0.007 = $0.5817
-
Total expected: $1.65
Actual Behavior
Dashboard shows:
-
DeepSeek V4 Flash: $1.96
-
web_search: $0.01
-
Total actual: $1.97
That's a $0.31 difference
Steps to reproduce the issue
- Use DeepSeek V4 Flash on the GOAT plan during off-peak hours.
- Generate significant Cache Read usage (83.1M tokens over 2 days).
- Compare the actual billed amount against the official pricing table.
Supporting Information
- Official pricing table: https://commandcode.ai/docs/plans/goat
- Usage breakdown (2 days):
- Cache Read: 83.1M
- Input (uncached): 2.5M
- Output: 0.7874M
- Total tokens: 86.4M
Command Code Version
1.29.0
Environment
- Plan: GOAT
- Model: DeepSeek V4 Flash
- Usage Period: August 17–18, 2026
- Usage Window: All off-peak hours (no peak-time surcharges should apply)
- ZDR (Zero Data Retention): Not enabled
Operating System
macOS
Terminal/IDE
WezTerm
Shell
zsh
Session file (optional)
No response
Fix prompt (optional)
No response
Additional context
I contacted support and was told that the GOAT plan uses credits and each model consumes them at a different rate. However, this doesn't explain why only Cache Read deviates from the table while Input and Output match perfectly.
I've also discussed this with other users and the hourly pricing theory was suggested (peak/off-peak rates), but since all my usage was during off-peak hours, that doesn't account for the discrepancy.
Note on Trace IDs:
I cannot access the individual trace IDs for those days because the Command Code Studio UI only displays the last 100 requests. I've asked support if there's a way to retrieve the full history, but I don't have the granular per-request data to pinpoint exactly which requests caused the overage. If there's an API endpoint or log export I can use to get the full trace history, please let me know.
- Vorherrschende Sprache
- Keine Sprachdaten
- Sterne
- 4k
- Forks
- 350
- PR-Merge-Kennzahlen
- Keine gemergten PRs in 30 T.
Beitragsleitfaden
Für dieses Repository ist kein Beitragsleitfaden indexiert
Erste Schritte
- Lesen Sie das ganze Issue und danach den Beitragsleitfaden des Projekts.
- Schreiben Sie ins Issue, dass Sie es übernehmen — das erspart doppelte Arbeit.
- Forken Sie das Repository und arbeiten Sie in einem Branch.
- Öffnen Sie einen Pull Request, der die Issue-Nummer nennt.
Mehr aus CommandCodeAI/command-code
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
CommandCodeAI/command-code#855 ·
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 78/100
CommandCodeAI/command-code#841 · 1 Kommentar ·
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
CommandCodeAI/command-code#655 · 1 Kommentar ·
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 68/100
CommandCodeAI/command-code#608 ·
-
Schwierigkeit 3/5 1-2 Tage Anfängerfreundlichkeit 70/100
CommandCodeAI/command-code#893 ·
Alle Issues in CommandCodeAI/command-code
Ähnliche Issues
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 78/100
use-agent-os/agent-os#3263 ·
-
[Bug]: context-limit error parsing has no pattern for llama.cpp's "context size (N tokens)" phrasing Offenarea/compression area/local-models area/sessions comp/agent duplicate P2 sweeper:risk-session-state type/bug
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 82/100
NousResearch/hermes-agent#117793 · 1 Kommentar ·
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 82/100
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 84/100
-
Schwierigkeit 2/5 1-3 Stunden Anfängerfreundlichkeit 75/100
BasedHardware/omi#15236 · 1 Kommentar ·