CommandCodeAI / CommandCodeAI/command-code

GOAT Plan: DeepSeek V4 Flash Cache Read billed ~53% higher than published rate (off-peak usage)

オープン
#722 コメント 3 件 リアクション 0 件 担当者 0 名 GitHub で見る

まだ誰も着手していません。

主要言語
言語のデータがありません
スター
4k
フォーク
350
PR マージ指標
30日以内にマージされた PR はありません

説明

Summary
Description

I've identified a billing discrepancy on the GOAT plan when using DeepSeek V4 Flash. The actual cost for Cache Read tokens is significantly higher than the published rate in the official pricing table, while Input and Output costs match the table exactly.

Expected Behavior

Based on the official GOAT pricing table for DeepSeek V4 Flash:

  • Input: $0.22/M
  • Output: $0.66/M
  • Cache Read: $0.007/M

With my usage (all during off-peak hours):

  • Input: 2.5M × $0.22 = $0.55

  • Output: 0.7874M × $0.66 = $0.5197

  • Cache Read: 83.1M × $0.007 = $0.5817

  • Total expected: $1.65

Actual Behavior

Dashboard shows:

  • DeepSeek V4 Flash: $1.96

  • web_search: $0.01

  • Total actual: $1.97

That's a $0.31 difference

Image Image
Steps to reproduce the issue
  1. Use DeepSeek V4 Flash on the GOAT plan during off-peak hours.
  2. Generate significant Cache Read usage (83.1M tokens over 2 days).
  3. Compare the actual billed amount against the official pricing table.
Supporting Information
Command Code Version

1.29.0

Environment
  • Plan: GOAT
  • Model: DeepSeek V4 Flash
  • Usage Period: August 17–18, 2026
  • Usage Window: All off-peak hours (no peak-time surcharges should apply)
  • ZDR (Zero Data Retention): Not enabled
Operating System

macOS

Terminal/IDE

WezTerm

Shell

zsh

Session file (optional)

No response

Fix prompt (optional)

No response

Additional context

I contacted support and was told that the GOAT plan uses credits and each model consumes them at a different rate. However, this doesn't explain why only Cache Read deviates from the table while Input and Output match perfectly.

I've also discussed this with other users and the hourly pricing theory was suggested (peak/off-peak rates), but since all my usage was during off-peak hours, that doesn't account for the discrepancy.

Note on Trace IDs:

I cannot access the individual trace IDs for those days because the Command Code Studio UI only displays the last 100 requests. I've asked support if there's a way to retrieve the full history, but I don't have the granular per-request data to pinpoint exactly which requests caused the overage. If there's an API endpoint or log export I can use to get the full trace history, please let me know.

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

はじめの一歩

  1. issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
  2. 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
  3. リポジトリをフォークし、ブランチを切って変更します。
  4. issue 番号を参照したプルリクエストを送ります。

調査の方向性

まず GOAT の料金表と DeepSeek V4 Flash のダッシュボード上の使用量内訳を確認し、その後、完全なトレース履歴が API またはログエクスポートで利用できるかを調査します。報告されたオフピーク期間中の Cache Read の課金を公開料金と比較します。不一致が修正されるか、その課金基準が明確に説明されれば完了です。

索引モデルが issue の本文から書いたものです。

評価

領域
ai, payments
issue の種類
バグ
難易度
4/5
見積もり時間
3〜5日
活発さ
活発
明瞭さ
説明が足りない
初心者へのやさしさ
35/100

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。