← all tools

LLM API Pricing Calculator

Live pricing for hundreds of models, straight from OpenRouter. Output tokens cost far more than input, so enter your monthly usage to see the real cost per model updated automatically.

Live prices via OpenRouter, refreshed on every visit · per 1M tokens; batch ~50% off, prompt caching up to 90% off input · confirm on the provider

usefulHQ Study · 2026

The same month of AI costs $0.2783 on DeepSeek vs $50 on Claude Opus, a 180× spread

We priced a typical assistant workload, 5 million input and 1 million output tokens a month, across the flagship models using live OpenRouter rates. The cheapest capable option, DeepSeek V4 Flash 0423, runs about $0.2783 a month; the priciest, Claude Opus 5, about $50. For most tasks a mid-tier model does the job at a fraction of the frontier price, so matching the model to the task, not defaulting to the biggest name, is where the savings are.

Cheapest flagship
$0.2783/mo
DeepSeek
Priciest flagship
$50/mo
Claude Opus
Cost spread
180×
same workload
Monthly cost by model (5M in + 1M out)
DeepSeek$0.2783GPT-5$2.2Gemini Flash$4Grok$7Gemini Pro$16.25Claude Sonnet$20Claude Opus$50

Method: live OpenRouter per-token rates × 5M input + 1M output tokens/month; batch and prompt-caching discounts excluded. Cite this: "usefulHQ LLM Cost Study 2026", tools.usefulhq.com/llm-pricing/. Updates from live prices. More usefulHQ studies →

Embed or cite this stat
Paste this into your page (a free link back to the data):
million
million
Your cost/mo ▲ModelProviderInput /1MOutput /1MContext
FreeLing 3.0 Flash VL (free)👁🧠inclusionAI262KOpenRouter →
FreeNex-N2.5-Mini (free)👁🧠Nex AGI262KOpenRouter →
FreeNex-N2.5-Pro (free)👁🧠Nex AGI262KOpenRouter →
FreeLing 3.0 Flash Sante (free)🧠inclusionAI262KOpenRouter →
FreeLing 3.0 Flash Fin (free)🧠inclusionAI262KOpenRouter →
FreeQwen3.8 27B (free)👁🧠Qwen262KOpenRouter →
FreeDots3-Note Preview (free)👁🧠Dots Studio512KOpenRouter →
FreeLFM2.5-2.6B (free)🧠LiquidAI66KOpenRouter →
FreeNemotron 3.5 Lightning (free)🧠NVIDIA1MOpenRouter →
FreeDeepSeek V4 Flash 0731 (free)🧠DeepSeek1.048576MOpenRouter →
FreeInkling Small (free)👁🧠Thinking Machines1.048576MOpenRouter →
FreeLaguna S 2.1 (free)🧠Poolside262KOpenRouter →
FreeInkling (free)👁🧠Thinking Machines1.048576MOpenRouter →
FreeLaguna XS 2.1 (free)🧠Poolside262KOpenRouter →
FreeNorth Mini Code (free)🧠Cohere256KOpenRouter →
FreeGLM 5.2 (free)🧠Z.ai33KOpenRouter →
FreeNemotron 3.5 Content Safety (free)👁🧠NVIDIA128KOpenRouter →
FreeNemotron 3 Ultra (free)🧠NVIDIA1MOpenRouter →
FreeNemotron 3 Nano Omni (free)👁🧠NVIDIA256KOpenRouter →
FreeGemma 4 26B A4B (free)👁🧠Google262KOpenRouter →
FreeGemma 4 31B (free)👁🧠Google262KOpenRouter →
FreeLyria 3 Pro Preview👁Google1.048576MOpenRouter →
FreeLyria 3 Clip Preview👁Google1.048576MOpenRouter →
FreeNemotron 3 Super (free)🧠NVIDIA262KOpenRouter →
FreeFree Models Router👁🧠Other200KOpenRouter →
$0.125CHEAPESTMistral NemoMistral$0.019$0.03131KOpenRouter →
$0.168Ling 3.0 Flash🧠inclusionAI$0.021$0.063262KOpenRouter →
$0.197Granite 4.0 MicroIBM$0.017$0.112131KOpenRouter →
$0.25Llama 3 8B LunarisSao10K$0.04$0.058KOpenRouter →
$0.2783DeepSeek V4 Flash 0423🧠DeepSeek$0.0398$0.07951.048576MOpenRouter →
$0.28DeepSeek V4 Flash Latest🧠DeepSeek$0.04$0.081.31072MOpenRouter →
$0.28DeepSeek V4 Flash 0731🧠DeepSeek$0.04$0.081.31072MOpenRouter →
$0.28Qwen3.7 Flash👁🧠Qwen$0.03$0.131MOpenRouter →
$0.28gpt-oss-20b🧠OpenAI$0.03$0.13131KOpenRouter →
$0.3Schematron V2 TurboInference.net$0.03$0.15128KOpenRouter →
$0.315Nova Micro 1.0Amazon$0.035$0.14128KOpenRouter →
$0.325GPT-5 Nano (batch)👁🧠OpenAI$0.025$0.2400KOpenRouter →
$0.33Mistral Small 3Mistral$0.05$0.0833KOpenRouter →
$0.33Llama 3.1 8B InstructMeta$0.05$0.08131KOpenRouter →
$0.336Llama 3.2 1B InstructMeta$0.027$0.20160KOpenRouter →
$0.3375Command R7B (12-2024)Cohere$0.0375$0.15128KOpenRouter →
$0.35Mercury 2.5🧠Inception$0.04$0.15260KOpenRouter →
$0.35Gemma 3 4B👁Google$0.05$0.1131KOpenRouter →
$0.397Hy-MT2-1.8BTencent$0.044$0.1778KOpenRouter →
$0.4Gemma 3 12B👁Google$0.05$0.15131KOpenRouter →
$0.42Laguna XS 2.1🧠Poolside$0.06$0.12262KOpenRouter →
$0.4338Qwen3 30B A3B Instruct 2507Qwen$0.0482$0.1931262KOpenRouter →
$0.45Gemini 2.5 Flash Lite (batch)👁🧠Google$0.05$0.21.048576MOpenRouter →
$0.45GPT-4.1 Nano (batch)👁OpenAI$0.05$0.21.047576MOpenRouter →
$0.45Ministral 3 8B 2512 (batch)👁Mistral$0.075$0.075262KOpenRouter →
$0.48Schematron V2 SmallInference.net$0.05$0.23128KOpenRouter →
$0.48Ling 3.0 Flash VL👁🧠inclusionAI$0.06$0.18131KOpenRouter →
$0.48Ling 3.0 Flash Fin🧠inclusionAI$0.06$0.18262KOpenRouter →
$0.49Phi 4Microsoft$0.07$0.1416KOpenRouter →
$0.51MythoMax 13BOther$0.08$0.118KOpenRouter →
$0.54Nemotron 3 Nano 30B A3B🧠NVIDIA$0.06$0.24262KOpenRouter →
$0.54Nova Lite 1.0👁Amazon$0.06$0.24300KOpenRouter →
$0.55Granite 4.2 8B🧠IBM$0.06$0.25131KOpenRouter →
$0.55Nemotron 3.5 Lightning🧠NVIDIA$0.07$0.2262KOpenRouter →
$0.58Llama 3.2 3B InstructMeta$0.05$0.33131KOpenRouter →
$0.585Qwen3.5-Flash👁🧠Qwen$0.065$0.261MOpenRouter →
$0.6Reka Edge👁🧠Other$0.1$0.116KOpenRouter →
$0.6Ministral 3 3B 2512👁Mistral$0.1$0.1131KOpenRouter →
$0.625GLM Flash Latest👁🧠Z.ai$0.075$0.251.31072MOpenRouter →
$0.625GLM 5.3 Flash (batch)👁🧠Z.ai$0.075$0.251.048576MOpenRouter →
$0.63Laguna S 2.1🧠Poolside$0.09$0.181.048576MOpenRouter →
$0.63Qwen3 Coder 30B A3B InstructQwen$0.07$0.28262KOpenRouter →
$0.65Qwen3.5-9B👁🧠Qwen$0.1$0.15262KOpenRouter →
$0.65GPT-5 Nano👁🧠OpenAI$0.05$0.4400KOpenRouter →
$0.665Hy-MT2-30B-A3BTencent$0.074$0.2958KOpenRouter →
$0.665Hy-MT2-7BTencent$0.074$0.2958KOpenRouter →
$0.675Mistral Small 4 (batch)👁🧠Mistral$0.075$0.3262KOpenRouter →
$0.675Seed 1.6 Flash👁🧠ByteDance Seed$0.075$0.3262KOpenRouter →
$0.675gpt-oss-safeguard-20b🧠OpenAI$0.075$0.3131KOpenRouter →
$0.675GPT-4o-mini (batch)👁OpenAI$0.075$0.3128KOpenRouter →
$0.68Qwen3 32B🧠Qwen$0.08$0.28131KOpenRouter →
$0.7Muse Spark 1.3 Contributor👁🧠Meta$0.1$0.21.048576MOpenRouter →
$0.7Muse Spark 1.2 Contributor👁🧠Meta$0.1$0.21.048576MOpenRouter →
$0.7UI-TARS 7B 👁ByteDance$0.1$0.2128KOpenRouter →
$0.7Reka Flash 3🧠Other$0.1$0.266KOpenRouter →
$0.7Qwen2.5 7B InstructQwen$0.1$0.233KOpenRouter →
$0.7025GLM 4.7 Flash🧠Z.ai$0.0605$0.4200KOpenRouter →
$0.7188Mistral Small 3.2 24B👁Mistral$0.0938$0.25256KOpenRouter →
$0.7425Hy3🧠Tencent$0.0825$0.33262KOpenRouter →
$0.75GLM 5.3 Flash👁🧠Z.ai$0.09$0.31.31072MOpenRouter →
$0.75Gemma 4 26B A4B 👁🧠Google$0.09$0.3262KOpenRouter →
$0.7875Qwen3 235B A22B Instruct 2507Qwen$0.0875$0.35262KOpenRouter →
$0.79Gemma 4 31B👁🧠Google$0.09$0.34262KOpenRouter →
$0.8Step 3.5 Flash🧠StepFun$0.1$0.3262KOpenRouter →
$0.8Voxtral Small 24B 2507Mistral$0.1$0.333KOpenRouter →
$0.8Llama 4 Scout👁Meta$0.1$0.31.31072MOpenRouter →
$0.81Solar Pro 4🧠Upstage$0.09$0.36524KOpenRouter →
$0.82Llama 3.3 70B InstructMeta$0.1$0.32131KOpenRouter →
$0.84Qwen3 14B🧠Qwen$0.12$0.24131KOpenRouter →
$0.85Nemotron 3 Super🧠NVIDIA$0.08$0.45262KOpenRouter →
$0.85Gemma 3 27B👁Google$0.08$0.45131KOpenRouter →
$0.875Ternary Bonsai 2 27B👁🧠PrismML$0.075$0.5262KOpenRouter →
$0.88DeepSeek V4 Flash Vision Exp (batch)👁🧠DeepSeek$0.11$0.331.048576MOpenRouter →
$0.88DeepSeek V4 Flash 0731 (batch)🧠DeepSeek$0.11$0.331.048576MOpenRouter →
$0.9Seed-2.0-Mini👁🧠ByteDance Seed$0.1$0.4262KOpenRouter →
$0.9Gemini 2.5 Flash Lite👁🧠Google$0.1$0.41.048576MOpenRouter →
$0.9GPT-4.1 Nano👁OpenAI$0.1$0.41.047576MOpenRouter →
$0.9Ministral 3 8B 2512👁Mistral$0.15$0.15262KOpenRouter →
$0.936Qwen3 VL 32B Instruct👁Qwen$0.104$0.416131KOpenRouter →
$0.98MiMo-V2.5👁🧠Xiaomi$0.14$0.281.05MOpenRouter →
$1.04Qwen3 VL 8B Instruct👁Qwen$0.117$0.455262KOpenRouter →
$1.04Qwen3 8B🧠Qwen$0.117$0.455131KOpenRouter →
$1.08Llama Guard 4 12B👁Meta$0.18$0.18164KOpenRouter →
$1.1GPT-5.6 Luna Pro (batch)👁🧠OpenAI$0.1$0.61.05MOpenRouter →
$1.1GPT-5.6 Luna (batch)👁🧠OpenAI$0.1$0.61.05MOpenRouter →
$1.1Qwen3.5-9B (batch)👁🧠Qwen$0.17$0.25262KOpenRouter →
$1.1Qwen3 30B A3B🧠Qwen$0.12$0.5131KOpenRouter →
$1.13GPT-5.4 Nano (batch)👁🧠OpenAI$0.1$0.625400KOpenRouter →
$1.17DeepSeek Flash Latest👁🧠DeepSeek$0.13$0.521.048576MOpenRouter →
$1.17Qwen3 VL 30B A3B Instruct👁Qwen$0.13$0.52262KOpenRouter →
$1.2Nemotron 3.5 Content Safety👁🧠NVIDIA$0.2$0.2131KOpenRouter →
$1.2Ministral 3 14B 2512👁Mistral$0.2$0.2262KOpenRouter →
$1.2Codestral 2508 (batch)Mistral$0.15$0.45256KOpenRouter →
$1.22Qwen3.8 Flash👁🧠Qwen$0.15$0.471MOpenRouter →
$1.27Hunyuan A13B Instruct🧠Tencent$0.14$0.57131KOpenRouter →
$1.35DeepSeek V4.1 Flash👁🧠DeepSeek$0.15$0.61.048576MOpenRouter →
$1.35Mistral Small 4👁🧠Mistral$0.15$0.6262KOpenRouter →
$1.35Solar Pro 3🧠Upstage$0.15$0.6131KOpenRouter →
$1.35gpt-oss-120b🧠OpenAI$0.15$0.6131KOpenRouter →
$1.35gpt-oss-120b (batch)🧠OpenAI$0.15$0.6131KOpenRouter →
$1.35Command R (08-2024)Cohere$0.15$0.6128KOpenRouter →
$1.35GPT-4o-mini👁OpenAI$0.15$0.6128KOpenRouter →
$1.35GPT-4o-mini (2024-07-18)👁OpenAI$0.15$0.6128KOpenRouter →
$1.38Gemini 3.1 Flash Lite (batch)👁🧠Google$0.125$0.751.048576MOpenRouter →
$1.4Qwen3.6 35B A3B👁🧠Qwen$0.1$0.9262KOpenRouter →
$1.4Qwen3 Coder NextQwen$0.12$0.8262KOpenRouter →
$1.5Hy3 preview🧠Tencent$0.18$0.6262KOpenRouter →
$1.5GLM 4.5 Air🧠Z.ai$0.13$0.85131KOpenRouter →
$1.55Qwen3 Next 80B A3B InstructQwen$0.09$1.1262KOpenRouter →
$1.59Llama 4 Maverick👁Meta$0.1875$0.65251.048576MOpenRouter →
$1.6SabaMistral$0.2$0.633KOpenRouter →
$1.63Muse Glimmer 30B (batch)👁🧠Meta$0.175$0.75131KOpenRouter →
$1.63GPT-5 Mini (batch)👁🧠OpenAI$0.125$1400KOpenRouter →
$1.72DeepSeek V4 Flash Vision Exp👁🧠DeepSeek$0.2156$0.64681.048576MOpenRouter →
$1.74DeepSeek V3.2🧠DeepSeek$0.269$0.4164KOpenRouter →
$1.76DeepSeek V3.2 Exp🧠DeepSeek$0.27$0.41164KOpenRouter →
$1.8GPT-4.1 Mini (batch)👁OpenAI$0.2$0.81.047576MOpenRouter →
$1.9UncensoredVenice$0.2$0.9128KOpenRouter →
$1.95Qwen3 Next 80B A3B Thinking🧠Qwen$0.15$1.2262KOpenRouter →
$1.95Qwen3 Coder FlashQwen$0.195$0.9751MOpenRouter →
$2Gemini 3.5 Flash Lite (batch)👁🧠Google$0.15$1.251.048576MOpenRouter →
$2Mercury 2🧠Inception$0.25$0.75128KOpenRouter →
$2Mistral Large 3 2512 (batch)👁Mistral$0.25$0.75262KOpenRouter →
$2Cydonia 24B V4.1TheDrummer$0.3$0.5131KOpenRouter →
$2Mistral Medium 3.1 (batch)👁Mistral$0.2$1131KOpenRouter →
$2Gemini 2.5 Flash (batch)👁🧠Google$0.15$1.251.048576MOpenRouter →
$2GPT-3.5 Turbo (batch)OpenAI$0.25$0.7516KOpenRouter →
$2.05Trinity Large Thinking🧠Arcee AI$0.25$0.8262KOpenRouter →
$2.06Qwen3.6 Flash👁🧠Qwen$0.1875$1.131MOpenRouter →
$2.08Qwen Plus 0728🧠Qwen$0.26$0.781MOpenRouter →
$2.08Qwen-PlusQwen$0.26$0.781MOpenRouter →
$2.1MiniMax-01👁MiniMax$0.2$1.11.000192MOpenRouter →
$2.11Qwen3.5-35B-A3B👁🧠Qwen$0.1625$1.3262KOpenRouter →
$2.15Step 3.7 Flash👁🧠StepFun$0.2$1.15262KOpenRouter →
$2.2GPT Luna Latest👁🧠OpenAI$0.2$1.21.05MOpenRouter →
How we rank. We cost your actual workload: input tokens × the input rate plus output tokens × the output rate. That matters because output usually costs 3-6× more than input so a model that looks cheap on input can be dear if you generate a lot. "Blended /1M" assumes a typical 3:1 input:output mix for a quick single-number comparison. All rates are standard USD per million tokens, converted live to your currency, batch APIs are ~50% cheaper and prompt caching can cut cached input up to 90%. Prices move fast; confirm on the provider's site.

Common questions

What's the cheapest LLM API?

Proprietary: Gemini Flash-Lite and GPT-4.1 Nano near $0.10/1M input. Cheapest overall is DeepSeek (~$0.14 in / $0.28 out). The best pick depends on your input:output ratio, use the calculator.

Why do input and output cost different amounts?

Generating output is far more expensive than reading input, so output is usually 3-6× the input price. That's why a single blended price is misleading.

Is Claude or GPT cheaper?

Close at the frontier. Claude Opus ~$5/$25, GPT-5 ~$2.50/$15 standard (much more for Pro). Claude's prompt caching can cut input up to 90%, often making it cheaper in practice.

How do I cut LLM costs?

Use a smaller model where you can (budget models are 20-50× cheaper), cache repeated prompts, use the batch API for non-urgent jobs, and trim prompt and output length.

The other way to pay for tokens

Per-million-token pricing only makes sense against the alternative, which is buying the hardware once and running the model yourself. These track what that hardware costs.

Related calculators & tools
Electricity cost →Fuel cost →Car recalls →AU PR points →