CommandCodeAI / CommandCodeAI/command-code

GOAT Plan: DeepSeek V4 Flash Cache Read billed ~53% higher than published rate (off-peak usage)

Aperta
#722 3 commenti 0 reazioni 0 assegnatari Vedi su GitHub

Nessuno ha ancora preso questa issue.

Lingua principale
Nessun dato sulla lingua
Stelle
4k
Fork
350
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Summary
Description

I've identified a billing discrepancy on the GOAT plan when using DeepSeek V4 Flash. The actual cost for Cache Read tokens is significantly higher than the published rate in the official pricing table, while Input and Output costs match the table exactly.

Expected Behavior

Based on the official GOAT pricing table for DeepSeek V4 Flash:

  • Input: $0.22/M
  • Output: $0.66/M
  • Cache Read: $0.007/M

With my usage (all during off-peak hours):

  • Input: 2.5M × $0.22 = $0.55

  • Output: 0.7874M × $0.66 = $0.5197

  • Cache Read: 83.1M × $0.007 = $0.5817

  • Total expected: $1.65

Actual Behavior

Dashboard shows:

  • DeepSeek V4 Flash: $1.96

  • web_search: $0.01

  • Total actual: $1.97

That's a $0.31 difference

Image Image
Steps to reproduce the issue
  1. Use DeepSeek V4 Flash on the GOAT plan during off-peak hours.
  2. Generate significant Cache Read usage (83.1M tokens over 2 days).
  3. Compare the actual billed amount against the official pricing table.
Supporting Information
Command Code Version

1.29.0

Environment
  • Plan: GOAT
  • Model: DeepSeek V4 Flash
  • Usage Period: August 17–18, 2026
  • Usage Window: All off-peak hours (no peak-time surcharges should apply)
  • ZDR (Zero Data Retention): Not enabled
Operating System

macOS

Terminal/IDE

WezTerm

Shell

zsh

Session file (optional)

No response

Fix prompt (optional)

No response

Additional context

I contacted support and was told that the GOAT plan uses credits and each model consumes them at a different rate. However, this doesn't explain why only Cache Read deviates from the table while Input and Output match perfectly.

I've also discussed this with other users and the hourly pricing theory was suggested (peak/off-peak rates), but since all my usage was during off-peak hours, that doesn't account for the discrepancy.

Note on Trace IDs:

I cannot access the individual trace IDs for those days because the Command Code Studio UI only displays the last 100 requests. I've asked support if there's a way to retrieve the full history, but I don't have the granular per-request data to pinpoint exactly which requests caused the overage. If there's an API endpoint or log export I can use to get the full trace history, please let me know.

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Come iniziare

  1. Leggi tutta la issue e poi la guida ai contributi del progetto.
  2. Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
  3. Fai un fork del repository e lavora su un branch.
  4. Apri una pull request che faccia riferimento al numero della issue.

Direzione di ricerca

Inizia con la tabella dei prezzi di GOAT e la suddivisione dell’utilizzo nella dashboard per DeepSeek V4 Flash, quindi verifica se la cronologia completa delle tracce è disponibile tramite un’API o un’esportazione dei log. Confronta la fatturazione di Cache Read durante il periodo di bassa affluenza segnalato con la tariffa pubblicata; il lavoro è completato quando la discrepanza viene corretta oppure la sua base di fatturazione viene spiegata chiaramente.

Scritto dal modello di indicizzazione a partire dal testo della issue.

Valutazione

Ambito
ai, payments
Tipo di issue
Bug
Difficoltà
4/5
Tempo stimato
3-5 giorni
Stato di attività
Attiva
Chiarezza
Da chiarire
Idoneità per principianti
35/100

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.