CommandCodeAI / CommandCodeAI/command-code
GOAT Plan: DeepSeek V4 Flash Cache Read billed ~53% higher than published rate (off-peak usage)
Personne n'a encore pris cette issue.
- Langage dominant
- Aucune donnée de langage
- Étoiles
- 4k
- Forks
- 350
- Métriques de merge des PR
- Aucune PR mergée en 30 j
Description
Summary
Description
I've identified a billing discrepancy on the GOAT plan when using DeepSeek V4 Flash. The actual cost for Cache Read tokens is significantly higher than the published rate in the official pricing table, while Input and Output costs match the table exactly.
Expected Behavior
Based on the official GOAT pricing table for DeepSeek V4 Flash:
- Input: $0.22/M
- Output: $0.66/M
- Cache Read: $0.007/M
With my usage (all during off-peak hours):
-
Input: 2.5M × $0.22 = $0.55
-
Output: 0.7874M × $0.66 = $0.5197
-
Cache Read: 83.1M × $0.007 = $0.5817
-
Total expected: $1.65
Actual Behavior
Dashboard shows:
-
DeepSeek V4 Flash: $1.96
-
web_search: $0.01
-
Total actual: $1.97
That's a $0.31 difference
Steps to reproduce the issue
- Use DeepSeek V4 Flash on the GOAT plan during off-peak hours.
- Generate significant Cache Read usage (83.1M tokens over 2 days).
- Compare the actual billed amount against the official pricing table.
Supporting Information
- Official pricing table: https://commandcode.ai/docs/plans/goat
- Usage breakdown (2 days):
- Cache Read: 83.1M
- Input (uncached): 2.5M
- Output: 0.7874M
- Total tokens: 86.4M
Command Code Version
1.29.0
Environment
- Plan: GOAT
- Model: DeepSeek V4 Flash
- Usage Period: August 17–18, 2026
- Usage Window: All off-peak hours (no peak-time surcharges should apply)
- ZDR (Zero Data Retention): Not enabled
Operating System
macOS
Terminal/IDE
WezTerm
Shell
zsh
Session file (optional)
No response
Fix prompt (optional)
No response
Additional context
I contacted support and was told that the GOAT plan uses credits and each model consumes them at a different rate. However, this doesn't explain why only Cache Read deviates from the table while Input and Output match perfectly.
I've also discussed this with other users and the hourly pricing theory was suggested (peak/off-peak rates), but since all my usage was during off-peak hours, that doesn't account for the discrepancy.
Note on Trace IDs:
I cannot access the individual trace IDs for those days because the Command Code Studio UI only displays the last 100 requests. I've asked support if there's a way to retrieve the full history, but I don't have the granular per-request data to pinpoint exactly which requests caused the overage. If there's an API endpoint or log export I can use to get the full trace history, please let me know.
Guide de contribution
Aucun guide de contribution indexé pour ce dépôt
Par où commencer
- Lisez l'issue en entier, puis le guide de contribution du projet.
- Signalez en commentaire que vous la prenez — cela évite que deux personnes fassent le même travail.
- Forkez le dépôt et travaillez sur une branche.
- Ouvrez une pull request qui référence le numéro de l'issue.
Piste de recherche
Commencez par le tableau des tarifs de GOAT et la ventilation de l’utilisation dans le dashboard pour DeepSeek V4 Flash, puis vérifiez si l’historique complet des traces est disponible via une API ou un export de logs. Comparez la facturation de Cache Read pendant la période creuse signalée avec le tarif publié ; le travail est terminé lorsque l’écart est corrigé ou que sa base de facturation est clairement expliquée.
Rédigé par le modèle d'indexation à partir du texte de l'issue.
Évaluation
- Domaine
- ai, payments
- Type d'issue
- Bug
- Difficulté
- 4/5
- Temps estimé
- 3-5 jours
- Activité
- Active
- Clarté
- À clarifier
- Accessibilité débutants
- 35/100