← all tools

LLM API Pricing Calculator

Live pricing for hundreds of models, straight from OpenRouter. Output tokens cost far more than input, so enter your monthly usage to see the real cost per model updated automatically.

Live prices via OpenRouter, refreshed on every visit · per 1M tokens; batch ~50% off, prompt caching up to 90% off input · confirm on the provider

usefulHQ Study · 2026

The same month of AI costs $0.98 on DeepSeek vs $25 on Claude Opus, a 26× spread

We priced a typical assistant workload, 5 million input and 1 million output tokens a month, across the flagship models using live OpenRouter rates. The cheapest capable option, DeepSeek V4 Flash, runs about $0.98 a month; the priciest, Claude Opus 4.8 (batch), about $25. For most tasks a mid-tier model does the job at a fraction of the frontier price, so matching the model to the task, not defaulting to the biggest name, is where the savings are.

Cheapest flagship
$0.98/mo
DeepSeek
Priciest flagship
$25/mo
Claude Opus
Cost spread
26×
same workload
Monthly cost by model (5M in + 1M out)
DeepSeek$0.98Gemini Flash$2GPT-5$5.5Grok$7Gemini Pro$8.12Claude Sonnet$10Claude Opus$25

Method: live OpenRouter per-token rates × 5M input + 1M output tokens/month; batch and prompt-caching discounts excluded. Cite this: "usefulHQ LLM Cost Study 2026", usefulhq.com/llm-pricing/. Updates from live prices. More usefulHQ studies →

Embed or cite this stat
Paste this into your page (a free link back to the data):
million
million
Your cost/mo ▲ModelProviderInput /1MOutput /1MContext
FreeLing-3.0-flash (free)🧠Other262K
FreeLaguna S 2.1 (free)🧠Poolside262K
FreeLaguna XS 2.1 (free)🧠Poolside262K
FreeNorth Mini Code (free)🧠Cohere256K
FreeNemotron 3.5 Content Safety (free)👁🧠NVIDIA128K
FreeNemotron 3 Ultra (free)🧠NVIDIA1M
FreeNemotron 3 Nano Omni (free)👁🧠NVIDIA256K
FreeGemma 4 26B A4B (free)👁🧠Google262K
FreeGemma 4 31B (free)👁🧠Google262K
FreeLyria 3 Pro Preview👁Google1.04858M
FreeLyria 3 Clip Preview👁Google1.04858M
FreeNemotron 3 Super (free)🧠NVIDIA262K
FreeFree Models Router👁🧠Other200K
FreeNemotron 3 Nano 30B A3B (free)🧠NVIDIA256K
FreeNemotron Nano 12B 2 VL (free)👁🧠NVIDIA128K
FreeNemotron Nano 9B V2 (free)🧠NVIDIA128K
Freegpt-oss-20b (free)🧠OpenAI131K
$0.08Ling-2.6-flashinclusionAI$0.01$0.03262K
$0.125Mistral NemoMistral$0.019$0.03131K
$0.197Granite 4.0 MicroIBM$0.017$0.112131K
$0.225Nex-N2-Mini👁🧠Nex AGI$0.025$0.1262K
$0.25Llama 3 8B LunarisSao10K$0.04$0.058K
$0.28Qwen3.7 Flash👁🧠Qwen$0.03$0.131M
$0.28gpt-oss-20b🧠OpenAI$0.03$0.13131K
$0.315Nova Micro 1.0Amazon$0.035$0.14128K
$0.325GPT-5 Nano (batch)👁🧠OpenAI$0.025$0.2400K
$0.33Mistral Small 3Mistral$0.05$0.0833K
$0.33Llama 3.1 8B InstructMeta$0.05$0.08131K
$0.336Llama 3.2 1B InstructMeta$0.027$0.20160K
$0.3375Command R7B (12-2024)Cohere$0.0375$0.15128K
$0.35Granite 4.1 8BIBM$0.05$0.1131K
$0.35Gemma 3 4B👁Google$0.05$0.1131K
$0.355gpt-oss-120b🧠OpenAI$0.037$0.17131K
$0.36MythoMax 13BOther$0.06$0.068K
$0.4Gemma 3 12B👁Google$0.05$0.15131K
$0.42Laguna XS 2.1🧠Poolside$0.06$0.12262K
$0.42Gemma 3n 4BGoogle$0.06$0.1233K
$0.4338Qwen3 30B A3B Instruct 2507Qwen$0.0481$0.193262K
$0.45Nemotron 3 Nano 30B A3B🧠NVIDIA$0.05$0.2262K
$0.45Gemini 2.5 Flash Lite (batch)👁🧠Google$0.05$0.21.04858M
$0.49Phi 4Microsoft$0.07$0.1416K
$0.525Hy3 preview🧠Tencent$0.063$0.21262K
$0.54Nova Lite 1.0👁Amazon$0.06$0.24300K
$0.58Llama 3.2 3B InstructMeta$0.05$0.33131K
$0.585Qwen3.5-Flash👁🧠Qwen$0.065$0.261M
$0.6Reka Edge👁🧠Other$0.1$0.116K
$0.6Ministral 3 3B 2512👁Mistral$0.1$0.1131K
$0.62Qwen3 Coder 30B A3B InstructQwen$0.07$0.27262K
$0.65Qwen3.5-9B👁🧠Qwen$0.1$0.15262K
$0.65GPT-5 Nano👁🧠OpenAI$0.05$0.4400K
$0.675Seed 1.6 Flash👁🧠ByteDance Seed$0.075$0.3262K
$0.675gpt-oss-safeguard-20b🧠OpenAI$0.075$0.3131K
$0.68Qwen3 32B🧠Qwen$0.08$0.28131K
$0.69Gemma 4 26B A4B 👁🧠Google$0.07$0.34262K
$0.7Laguna S 2.1🧠Poolside$0.1$0.21.04858M
$0.7GLM 4.7 Flash🧠Z.ai$0.06$0.4203K
$0.7UI-TARS 7B 👁ByteDance$0.1$0.2128K
$0.7Reka Flash 3🧠Other$0.1$0.266K
$0.7Qwen2.5 7B InstructQwen$0.1$0.233K
$0.8Step 3.5 Flash🧠StepFun$0.1$0.3262K
$0.8Voxtral Small 24B 2507Mistral$0.1$0.332K
$0.8Mistral Small 3.2 24B👁Mistral$0.1$0.3256K
$0.8Llama 4 Scout👁Meta$0.1$0.31.31072M
$0.825Nemotron 3 Super🧠NVIDIA$0.085$0.41M
$0.84Gemma 4 31B👁🧠Google$0.1$0.34262K
$0.85Gemma 3 27B👁Google$0.08$0.45262K
$0.9Seed-2.0-Mini👁🧠ByteDance Seed$0.1$0.4262K
$0.9Gemini 2.5 Flash Lite👁🧠Google$0.1$0.41.04858M
$0.9GPT-4.1 Nano👁OpenAI$0.1$0.41.04758M
$0.9Ministral 3 8B 2512👁Mistral$0.15$0.15262K
$0.936Qwen3 VL 32B Instruct👁Qwen$0.104$0.416131K
$0.98DeepSeek V4 Flash🧠DeepSeek$0.14$0.281.04858M
$0.98MiMo-V2.5👁🧠Xiaomi$0.14$0.281.05M
$1Ring-2.6-1T🧠inclusionAI$0.075$0.625262K
$1Ling-2.6-1TinclusionAI$0.075$0.625262K
$1Qwen3 235B A22B Instruct 2507Qwen$0.09$0.55262K
$1.04Qwen3 VL 8B Instruct👁Qwen$0.117$0.455262K
$1.04Qwen3 8B🧠Qwen$0.117$0.455131K
$1.05Hermes 4 70B🧠Nous$0.13$0.4131K
$1.05Llama 3.3 70B InstructMeta$0.13$0.4131K
$1.08Llama Guard 4 12B👁Meta$0.18$0.181.04858M
$1.1Qwen3 30B A3B🧠Qwen$0.12$0.5131K
$1.12GPT-5.4 Nano (batch)👁🧠OpenAI$0.1$0.625400K
$1.19Hy3🧠Tencent$0.132$0.528262K
$1.2Ministral 3 14B 2512👁Mistral$0.2$0.2262K
$1.25Olmo 3 32B Think🧠AllenAI$0.15$0.566K
$1.27Hunyuan A13B Instruct🧠Tencent$0.14$0.57131K
$1.35KAT-Coder-Air V2.5Kwaipilot$0.15$0.6256K
$1.35MiniMax M3 (batch)👁🧠MiniMax$0.15$0.6524K
$1.35Mistral Small 4👁🧠Mistral$0.15$0.6262K
$1.35Solar Pro 3🧠Upstage$0.15$0.6128K
$1.35Qwen3 VL 30B A3B Instruct👁Qwen$0.15$0.6262K
$1.35Command R (08-2024)Cohere$0.15$0.6128K
$1.35GPT-4o-mini👁OpenAI$0.15$0.6128K
$1.35GPT-4o-mini (2024-07-18)👁OpenAI$0.15$0.6128K
$1.38Gemini 3.1 Flash Lite (batch)👁🧠Google$0.125$0.751.04858M
$1.4Qwen3 Coder NextQwen$0.12$0.8262K
$1.5GLM 4.5 Air🧠Z.ai$0.13$0.85131K
$1.6SabaMistral$0.2$0.633K
$1.6Qwen3 Next 80B A3B InstructQwen$0.1$1.1262K
$1.62GPT-5 Mini (batch)👁🧠OpenAI$0.125$1400K
$1.65MiniMax M2.5🧠MiniMax$0.15$0.9205K
$1.7Qwen3.6 35B A3B👁🧠Qwen$0.14$1262K
$1.7Qwen3.5-35B-A3B👁🧠Qwen$0.14$1262K
$1.74DeepSeek V3.2🧠DeepSeek$0.269$0.4164K
$1.75Rocinante 12BTheDrummer$0.25$0.566K
$1.76DeepSeek V3.2 Exp🧠DeepSeek$0.27$0.41164K
$1.8Llama 4 Maverick👁Meta$0.2$0.81.04858M
$1.9UncensoredVenice$0.2$0.9128K
$1.95Qwen3 Next 80B A3B Thinking🧠Qwen$0.15$1.2262K
$1.95Trinity Large Thinking🧠Arcee AI$0.22$0.85262K
$1.95Qwen3 Coder FlashQwen$0.195$0.9751M
$2Gemini 3.5 Flash Lite (batch)👁🧠Google$0.15$1.251.04858M
$2Mercury 2🧠Inception$0.25$0.75128K
$2Cydonia 24B V4.1TheDrummer$0.3$0.5131K
$2Gemini 2.5 Flash (batch)👁🧠Google$0.15$1.251.04858M
$2.05Qwen3 14B🧠Qwen$0.2275$0.91131K
$2.06Qwen3.6 Flash👁🧠Qwen$0.1875$1.121M
$2.08Qwen Plus 0728🧠Qwen$0.26$0.781M
$2.08Qwen-PlusQwen$0.26$0.781M
$2.1MiniMax-01👁MiniMax$0.2$1.11.00019M
$2.15Step 3.7 Flash👁🧠StepFun$0.2$1.15262K
$2.2Qwen2.5 72B InstructOther$0.36$0.433K
$2.2DeepSeek V3.1🧠DeepSeek$0.25$0.95164K
$2.25Nex-N2-Pro👁🧠Nex AGI$0.25$1262K
$2.25Perceptron Mk1👁🧠Perceptron$0.15$1.533K
$2.25MiniMax M2.7🧠MiniMax$0.25$1205K
$2.25GPT-5.4 Nano👁🧠OpenAI$0.2$1.25400K
$2.29MiniMax M2🧠MiniMax$0.255$1.02205K
$2.31Mistral Small 3.1 24B👁Mistral$0.351$0.555128K
$2.32DeepSeek V3DeepSeek$0.2574$1.03164K
$2.35DeepSeek V3.1 Terminus🧠DeepSeek$0.27$1164K
$2.4GLM 4.6V👁🧠Z.ai$0.3$0.9131K
$2.4Codestral 2508Mistral$0.3$0.9256K
$2.4UnslopNemo 12BTheDrummer$0.4$0.433K
$2.4Llama 3.1 70B InstructMeta$0.4$0.4131K
$2.47DeepSeek V3 0324DeepSeek$0.27$1.12164K
$2.5Qwen3 Coder 480B A35BQwen$0.3$1262K
$2.5Claude 3 Haiku👁Anthropic$0.25$1.25200K
$2.54Qwen3.5-27B👁🧠Qwen$0.195$1.56262K
$2.7LongCat 2.0🧠Meituan$0.3$1.21.04876M
$2.7MiniMax M3👁🧠MiniMax$0.3$1.21.04858M
$2.7KAT-Coder-Pro V2Kwaipilot$0.3$1.2262K
$2.7MiniMax M2-herMiniMax$0.3$1.266K
$2.7MiniMax M2.1🧠MiniMax$0.3$1.2205K
$2.75Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)👁🧠Google$0.25$1.566K
$2.75Gemini 3.1 Flash Lite👁🧠Google$0.25$1.51.04858M
$2.75Gemini 3.1 Flash Lite Preview👁🧠Google$0.25$1.51.04858M
$2.75Gemini 3 Flash Preview (batch)👁🧠Google$0.25$1.51.04858M
$2.86Qwen3.5 Plus 2026-02-15👁🧠Qwen$0.26$1.561M
$2.88Qwen3.7 Plus👁🧠Qwen$0.32$1.281M
$2.9ReMM SLERP 13BOther$0.45$0.656K
$2.95Qwen3 VL 235B A22B Instruct👁Qwen$0.21$1.9262K
$3Qwen3 VL 8B Thinking👁🧠Qwen$0.18$2.1131K
$3.04DeepSeek V4 Pro🧠DeepSeek$0.435$0.871.04858M
$3.04MiMo-V2.5-Pro🧠Xiaomi$0.435$0.871.05M
$3.2Qwen Plus 0728 (thinking)🧠Qwen$0.4$1.21M
$3.25Seed-2.0-Lite👁🧠ByteDance Seed$0.25$2262K
$3.25Seed 1.6👁🧠ByteDance Seed$0.25$2262K
$3.25GPT-5.1-Codex-Mini👁🧠OpenAI$0.25$2400K
How we rank. We cost your actual workload: input tokens × the input rate plus output tokens × the output rate. That matters because output usually costs 3-6× more than input so a model that looks cheap on input can be dear if you generate a lot. "Blended /1M" assumes a typical 3:1 input:output mix for a quick single-number comparison. All rates are standard USD per million tokens, converted live to your currency, batch APIs are ~50% cheaper and prompt caching can cut cached input up to 90%. Prices move fast; confirm on the provider's site.

Common questions

What's the cheapest LLM API?

Proprietary: Gemini Flash-Lite and GPT-4.1 Nano near $0.10/1M input. Cheapest overall is DeepSeek (~$0.14 in / $0.28 out). The best pick depends on your input:output ratio, use the calculator.

Why do input and output cost different amounts?

Generating output is far more expensive than reading input, so output is usually 3-6× the input price. That's why a single blended price is misleading.

Is Claude or GPT cheaper?

Close at the frontier. Claude Opus ~$5/$25, GPT-5 ~$2.50/$15 standard (much more for Pro). Claude's prompt caching can cut input up to 90%, often making it cheaper in practice.

How do I cut LLM costs?

Use a smaller model where you can (budget models are 20-50× cheaper), cache repeated prompts, use the batch API for non-urgent jobs, and trim prompt and output length.

Related calculators & tools
Electricity cost →Fuel cost →Car recalls →AU PR points →