LLM API budget planning

LLM API Cost Calculator

Estimate monthly API spend from input and output token volume, then compare popular and low-cost models using normalized USD prices per 1M tokens.

Page updated:  Data confirmed:  Prices normalized to USD per 1M tokens Calculator estimates planning cost, not final billing

Start with your workload

Change token volume once; every model estimate updates.

Use this when budgeting a chatbot, RAG pipeline, coding assistant, or batch analysis workload before choosing a provider.

Default scenario1M input + 500K output
Popular models12
Low-cost shortlist12
Comparable unitPer 1M tokens

Estimate your workload cost

Monthly token volume

Use input and output tokens separately; output-heavy apps can change the winner.

This estimate uses available model metadata. Provider invoices may include routing, caching, discounts, minimums, or account-specific terms.

Popular Model Cost Estimates

Open popular models
ModelProviderInput / 1MOutput / 1MYour CostContextPopularity
🔥DeepSeek V4 Flash 0731DeepSeek$0.08$0.18$0.171.31M#1
🔥MiMo-V2.5Xiaomi$0.14$0.28$0.281.05M#2
🔥Hy3Tencent$0.132$0.528$0.4262.14K#3
New🔥Ox Alphastealth$0$0$01.05M#4
🔥DeepSeek V4 Flash 0423DeepSeek$0.0517$0.1033$0.11.05M#5
🔥GPT-5.6 LunaOpenAI$0.2$1.2$0.81.05M#6
🔥Nemotron 3 Ultra (free)NVIDIA$0$0$01M#7
🔥GLM 5.2Z.ai$0.966$3.036$2.481.05M#8
🔥Claude Opus 5Anthropic$5$25$17.51M#9
🔥DeepSeek V4 Pro 0423DeepSeek$0.3969$0.7938$0.791.05M#10
🔥Laguna S 2.1 (free)Poolside$0$0$0262.14K#11
New🔥Gemini 3.7 FlashGoogle$0.375$1.875$1.311.05M#12

Low-Cost Shortlist

Browse cheapest models
ModelProviderYour CostInput / 1MOutput / 1MContext
New🔥Ox Alphastealth$0$0$01.05M
🔥Nemotron 3 Ultra (free)NVIDIA$0$0$01M
🔥Laguna S 2.1 (free)Poolside$0$0$0262.14K
New🔥Nemotron 3.5 Lightning (free)NVIDIA$0$0$01M
NewDots3-Note Preview (free)Dots Studio$0$0$0512K
NewLFM2.5-2.6B (free)LiquidAI$0$0$065.54K
Ling 3.0 Tiny (free)inclusionAI$0$0$0262.14K
Inkling Small (free)Thinking Machines$0$0$0262.14K
Ling-3.0-flash (free)inclusionAI$0$0$0262.14K
Inkling (free)Thinking Machines$0$0$0262.14K
Hy3 (free)Tencent$0$0$0262.14K
Laguna XS 2.1 (free)Poolside$0$0$0262.14K

Cost Planning FAQ

How does the calculator estimate cost?

It multiplies input tokens by the model's input price per 1M tokens, then adds output tokens multiplied by the model's output price per 1M tokens.

Why separate input and output tokens?

Chatbots, agents, and code assistants often spend more on output tokens, while retrieval and classification workloads may be input-heavy. Separating them prevents a cheap-looking model from winning the wrong workload.

What should I do after estimating cost?

Open the model page or compare it against a close alternative, then verify current provider limits, discounts, and availability before production use.