Explore AI
What it actually costs
Live prices for 446 models, translated into something you can picture: one long document summarised.
Prices fetched 2026-09-20 yesterday
“Task” = a 10,000-token document in, a 1,000-token answer out. That is one long PDF summarised, or a substantial code review.
| Model | Context | In /M | Out /M | One task |
|---|---|---|---|---|
| inclusionAI: Ling 3.0 Flash VL (free)inclusionai | 262k | free | free | free |
| Nex AGI: Nex-N2.5-Mini (free)nex-agi | 262k | free | free | free |
| Nex AGI: Nex-N2.5-Pro (free)nex-agi | 262k | free | free | free |
| inclusionAI: Ling 3.0 Flash Sante (free)inclusionai | 262k | free | free | free |
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai | 262k | free | free | free |
| Qwen: Qwen3.8 27B (free)qwen | 262k | free | free | free |
| Dots Studio: Dots3-Note Preview (free)dots-studio | 512k | free | free | free |
| LiquidAI: LFM2.5-2.6B (free)liquid | 66k | free | free | free |
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia | 1000k | free | free | free |
| Thinking Machines: Inkling Small (free)thinkingmachines | 1049k | free | free | free |
| Poolside: Laguna S 2.1 (free)poolside | 262k | free | free | free |
| Thinking Machines: Inkling (free)thinkingmachines | 1049k | free | free | free |
| Poolside: Laguna XS 2.1 (free)poolside | 262k | free | free | free |
| Cohere: North Mini Code (free)cohere | 256k | free | free | free |
| Z.ai: GLM 5.2 (free)z-ai | 33k | free | free | free |
| NVIDIA: Nemotron 3.5 Content Safety (free)nvidia | 128k | free | free | free |
| NVIDIA: Nemotron 3 Ultra (free)nvidia | 1000k | free | free | free |
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia | 256k | free | free | free |
| Google: Gemma 4 26B A4B (free)google | 262k | free | free | free |
| Google: Gemma 4 31B (free)google | 262k | free | free | free |
| Google: Lyria 3 Pro Previewgoogle | 1049k | free | free | free |
| Google: Lyria 3 Clip Previewgoogle | 1049k | free | free | free |
| NVIDIA: Nemotron 3 Super (free)nvidia | 262k | free | free | free |
| Free Models Routeropenrouter | 200k | free | free | free |
| IBM: Granite 4.0 Microibm-granite | 131k | $0.02 | $0.11 | $0.0003 |
| OpenAI: gpt-oss-20bopenai | 131k | $0.03 | $0.13 | $0.0004 |
| Sao10K: Llama 3 8B Lunarissao10k | 8k | $0.04 | $0.05 | $0.0004 |
| Google: Gemma 3 4Bgoogle | 131k | $0.05 | $0.10 | $0.0006 |
| inclusionAI: Ling 3.0 Flash Fininclusionai | 262k | $0.06 | $0.18 | $0.0008 |
| Qwen: Qwen3 Coder 30B A3B Instructqwen | 262k | $0.07 | $0.28 | $0.0010 |
| Mistral: Mistral Small 4 (batch)mistralai | 262k | $0.07 | $0.30 | $0.0010 |
| Google: Gemma 4 26B A4B google | 262k | $0.09 | $0.30 | $0.0012 |
| Qwen: Qwen2.5 7B Instructqwen | 33k | $0.10 | $0.20 | $0.0012 |
| Google: Gemma 3 27Bgoogle | 131k | $0.08 | $0.45 | $0.0013 |
| Mistral: Voxtral Small 24B 2507mistralai | 33k | $0.10 | $0.30 | $0.0013 |
| OpenAI: GPT-5.6 Luna (batch)openai | 1050k | $0.10 | $0.60 | $0.0016 |
| Xiaomi: MiMo-V2.5xiaomi | 1050k | $0.14 | $0.28 | $0.0017 |
| Qwen: Qwen3 30B A3Bqwen | 131k | $0.12 | $0.50 | $0.0017 |
| Mistral: Mistral Small 4mistralai | 262k | $0.15 | $0.60 | $0.0021 |
| Google: Gemini 2.5 Flash (batch)google | 1049k | $0.15 | $1.25 | $0.0027 |
| Inception: Mercury 2inception | 128k | $0.25 | $0.75 | $0.0033 |
| Qwen: Qwen Plus 0728qwen | 1000k | $0.26 | $0.78 | $0.0034 |
| Qwen: Qwen3 VL 8B Thinkingqwen | 131k | $0.18 | $2.10 | $0.0039 |
| Mistral: Codestral 2508mistralai | 256k | $0.30 | $0.90 | $0.0039 |
| Qwen: Qwen3 VL 235B A22B Instructqwen | 262k | $0.21 | $1.90 | $0.0040 |
| ReMM SLERP 13Bundi95 | 6k | $0.35 | $0.65 | $0.0041 |
| MiniMax: MiniMax M2.7minimax | 205k | $0.30 | $1.20 | $0.0042 |
| Qwen: Qwen3 VL 30B A3B Thinkingqwen | 262k | $0.20 | $2.40 | $0.0044 |
| OpenAI: GPT-5 Miniopenai | 400k | $0.25 | $2.00 | $0.0045 |
| Qwen: Qwen3.8 27Bqwen | 1000k | $0.20 | $2.55 | $0.0046 |
| Xiaomi: MiMo-V2.5-Proxiaomi | 1050k | $0.43 | $0.87 | $0.0052 |
| Google: Gemini 3.5 Flash Litegoogle | 1049k | $0.30 | $2.50 | $0.0055 |
| OpenAI: GPT-4.1 Miniopenai | 1048k | $0.40 | $1.60 | $0.0056 |
| OpenAI: GPT-5.4 Mini (batch)openai | 400k | $0.38 | $2.25 | $0.0060 |
| Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)google | 66k | $0.50 | $3.00 | $0.0080 |
| DeepSeek: DeepSeek V4 Pro 0813 (batch)deepseek | 1049k | $0.66 | $1.98 | $0.0086 |
| Qwen: Qwen3.5 397B A17Bqwen | 262k | $0.55 | $3.50 | $0.0090 |
| Sao10K: Llama 3.1 Euryale 70B v2.2sao10k | 131k | $0.85 | $0.85 | $0.0094 |
| AionLabs: Aion-RP 1.0 (8B)aion-labs | 33k | $0.80 | $1.60 | $0.0096 |
| MoonshotAI: Kimi K2.7 Codemoonshotai | 262k | $0.71 | $3.21 | $0.01 |
| OpenAI: GPT-5.1 (batch)openai | 400k | $0.63 | $5.00 | $0.01 |
| Writer: Palmyra X5writer | 1040k | $0.60 | $6.00 | $0.01 |
| OpenAI: GPT Mini Latest~openai | 400k | $0.75 | $4.50 | $0.01 |
| SpaceXAI: Grok Build 0.1x-ai | 256k | $1.00 | $2.00 | $0.01 |
| OpenAI: o3 (batch)openai | 200k | $1.00 | $4.00 | $0.01 |
| Thinking Machines: Inklingthinkingmachines | 1049k | $1.00 | $4.05 | $0.01 |
| OpenAI: o4 Miniopenai | 200k | $1.10 | $4.40 | $0.02 |
| Meta: Muse Spark 1.1meta | 1049k | $1.25 | $4.25 | $0.02 |
| OpenAI: GPT-5.1-Codexopenai | 400k | $1.25 | $10.00 | $0.02 |
| Mistral: Mistral Medium 3.5mistralai | 262k | $1.50 | $7.50 | $0.02 |
| Qwen: Qwen3.8 2.4T A95Bqwen | 1049k | $2.00 | $6.00 | $0.03 |
| SpaceXAI: Grok 4.5x-ai | 500k | $2.00 | $6.00 | $0.03 |
| Mistral: Mixtral 8x22B Instructmistralai | 66k | $2.00 | $6.00 | $0.03 |
| OpenAI: GPT-5.2-Codexopenai | 400k | $1.75 | $14.00 | $0.03 |
| Google: Gemini 3.1 Pro Previewgoogle | 1049k | $2.00 | $12.00 | $0.03 |
| OpenAI: GPT-4o (2024-11-20)openai | 128k | $2.50 | $10.00 | $0.04 |
| OpenAI: GPT-5.4openai | 1050k | $2.50 | $15.00 | $0.04 |
| Anthropic: Claude Sonnet 4.6anthropic | 1000k | $3.00 | $15.00 | $0.05 |
| OpenAI: GPT-4 Turbo (batch)openai | 128k | $5.00 | $15.00 | $0.07 |
| OpenAI: GPT-6 Astra (batch)openai | 1050k | $5.00 | $25.00 | $0.08 |
| OpenAI: GPT Chat Latestopenai | 400k | $5.00 | $30.00 | $0.08 |
| Anthropic: Claude Fable 5.1anthropic | 1000k | $10.00 | $50.00 | $0.15 |
| OpenAI: GPT-5.4 Pro (batch)openai | 1050k | $15.00 | $90.00 | $0.24 |
| OpenAI: GPT-5.5 Proopenai | 1050k | $30.00 | $180.00 | $0.48 |
The number that should change your mind
Input prices across the models on this page span about 8,824 times. That is not a small difference in quality for a small difference in price — it is a difference of three orders of magnitude, for models that all answer the same kind of question.
Output costs several times more than input on nearly every model. If you are generating long answers rather than reading long documents, your bill is dominated by output tokens. Ask for less, and pay less.
What actually drives your bill
- Output length, not input length, on most tasks. A terse instruction that produces a 200-token answer costs a fraction of a chatty one.
- Re-sending context. In a long conversation, the whole history is charged again on every turn. Long chats are quadratically expensive, which is why harnesses manage context carefully.
- Choosing a frontier model for a routine job. The cheapest paid models here are perfectly adequate for classification, extraction and formatting — tasks that make up most of the volume in real systems.
How to buy credits
Every route differs, and this is where people waste money:
- Directly from a lab — an account and a card per provider. Best price, worst convenience, one bill per vendor.
- An aggregator — one account and one credit balance reaching hundreds of models. Slightly more expensive per token, and worth it while you are deciding.
- Cloud platforms — the model is one line item on an existing cloud bill, which suits companies that already have one.
- Free tiers — real, and usually paid for with your data. Never send anything confidential through one.
- Prepaid credits — most providers let you buy credit rather than attach a card. Do that first: it caps the damage from a runaway loop.
Buy the smallest credit pack a provider offers and set a spend limit. An agent stuck in a retry loop can burn a surprising amount in an hour, and the invoice arrives later.