Skip to content
CostPerPrompt

What will AI actually cost you?

Live pricing for 232+ models, refreshed automatically — plus calculators that turn token prices into real answers: what your chatbot, agent, or API workload will cost per month.

Every cost question, answered

per 1M tokens · updated 2026-08-02
Model Vendor Input / 1M Output / 1M Context
GPT-5.6 Sol OpenAI $5.00 $30.00 1.1M
GPT-5.5 OpenAI $5.00 $30.00 1.1M
GPT-5.4 OpenAI $2.50 $15.00 1.1M
Claude Fable 5 Anthropic $10.00 $50.00 1M
Claude Opus 5 Anthropic $5.00 $25.00 1M
Claude Sonnet 5 Anthropic $2.00 $10.00 1M
Kimi K3 Moonshot (Kimi) $3.00 $15.00 1M
DeepSeek V4 Pro DeepSeek $0.435 $0.87 1M
DeepSeek V4 Flash 0731 DeepSeek $0.09 $0.18 1M
Gemini 3.1 Pro Preview Google $2.00 $12.00 1M
Gemini 3.6 Flash Google $1.50 $7.50 1M
Grok 4.5 xAI $2.00 $6.00 500K
Grok 4.3 xAI $1.25 $2.50 1M
GLM 5.2 Z.ai (GLM) $0.4186 $1.32 1M

See the full table of 232 models →

How AI API pricing works — the 60-second version

Every major AI provider bills the same way: you pay per token (roughly ¾ of a word), with separate rates for input (what you send) and output (what the model writes back). Output is usually 3–5× more expensive than input. A model listed at $5 / $25 per million tokens costs $5 for every million tokens you send and $25 for every million it generates.

Two discounts change the math dramatically: prompt caching cuts repeated input costs by up to 90% (critical for chatbots that resend conversation history), and batch processing takes ~50% off when you can wait for results. Our calculators account for both — most "how much will this cost" articles don't, which is why their estimates run 2–3× too high or too low.