Estimating Costs

Definition

Estimating costs means turning luv13 token counts into dollars with one multiplication, because every model has the same flat rate and input costs the same as output.

Key takeaways

  • Cost in USD = total tokens ÷ 1,000,000 × $0.33. There's no per-model table and no separate input and output rate.
  • 800,000 input + 200,000 output = 1,000,000 tokens = $0.33.
  • $5, the smallest top-up, covers about 15.15M tokens.
  • If you resend the whole conversation each turn, input grows every turn, and that's usually the biggest cost.
  • Failed or empty calls aren't charged. The rate is set on Pricing; if it changes, change the constant in your code.

Quick numbers

At $0.33 per 1M tokens:

TokensCost
1,000$0.00033
100,000$0.033
1,000,000$0.33
15,151,515about $5.00
RATE_PER_MILLION = 0.33  # USD, same for input and output

def cost_usd(total_tokens):
    return total_tokens / 1_000_000 * RATE_PER_MILLION

print(cost_usd(800_000 + 200_000))  # 0.33

Chats add up

If you send the full history each turn, and every turn adds 500 tokens of question and 500 of answer:

TurnInput sentOutputTokens this turn
15005001,000
21,5005002,000
32,5005003,000
109,50050010,000

Ten turns total 55,000 tokens, about $0.018, although only 10,000 tokens of new text were written. Trimming or summarizing old turns keeps this down; see Conversation History.

Your balance and actual usage are in the dashboard. See Usage and Billing.