Real-time pricing for 300+ AI models. Compare cost per million tokens, calculate your spend, find the cheapest model for your use case.
| Model | Provider | Input ($/1M) | Output ($/1M) | Context | Blended Cost * |
|---|
* Blended cost assumes 3:1 input-to-output ratio (typical chat workload). Prices are per 1 million tokens. Data sourced live from OpenRouter API. Some models may have additional fees (image processing, tool calls, caching).
LLM APIs charge per token — roughly ¾ of a word. A 1,000-word article is about 1,300 tokens. Most providers split pricing into input tokens (your prompt) and output tokens (the model's response). Output tokens typically cost 3-5× more than input.
Some providers offer batch pricing (50% off for non-real-time requests) and cached input discounts (up to 90% off for repeated prompt prefixes). Always check both input and output rates when comparing models — a cheap input price can hide expensive outputs.
This tracker pulls live pricing from OpenRouter, which aggregates 300+ models across all major providers. Prices update automatically.