Cutoff Data
AI model price tracker
What the major models actually cost per million tokens, derived from the billed cost of real API calls we make every hour — not from vendor rate cards.
Last reading 8/7/2026, 6:17:14 PM UTC
| Model | Provider | Input / 1M | Output / 1M | Context | Status |
|---|---|---|---|---|---|
| Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-lite | $0.10 | $0.40 | 1M | Responding | |
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | $0.25 | $1.50 | 1M | Responding | |
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | $0.30 | $2.50 | 1M | Responding | |
| Gemini 2.5 Progoogle/gemini-2.5-pro | $1.25 | $10.00 | 1M | Responding | |
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | $1.50 | $9.00 | 1M | Responding | |
| Gemini 3.1 Pro (preview)google/gemini-3.1-pro-preview | $2.00 | $12.00 | 1M | Responding | |
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | — | — | 1M | Responding | |
| GPT-5 Nanoopenai/gpt-5-nano | OpenAI | — | — | — | Responding |
| GPT-5 Miniopenai/gpt-5-mini | OpenAI | — | — | — | Responding |
| GPT-5.4 Miniopenai/gpt-5.4-mini | OpenAI | — | — | — | Responding |
| GPT-5.6 Solopenai/gpt-5.6-sol | OpenAI | — | — | — | Responding |
| GPT-5.6 Terraopenai/gpt-5.6-terra | OpenAI | — | — | — | Responding |
| GPT-5.6 Lunaopenai/gpt-5.6-luna | OpenAI | — | — | — | Responding |
How this is measured
Every hour we send an identical short prompt to each tracked model through a single gateway account and record what the call was billed. Dividing the billed input and output cost by the tokens consumed gives an observed price per million tokens.
Prices shown are therefore what one buyer was actually charged, at that hour, on that route. Enterprise agreements, batch discounts and cached-input rates will differ. Where a provider does not itemise cost, the price cell reads “—”.
A “not responding” model returned an error on the most recent sweep. That is a fact about our route to it, not necessarily a global outage.