Cutoff Data

AI model price tracker

What the major models actually cost per million tokens, derived from the billed cost of real API calls we make every hour — not from vendor rate cards.

Last reading 8/7/2026, 6:17:14 PM UTC

ModelProviderInput / 1MOutput / 1MContextStatus
Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-liteGoogle$0.10$0.401MResponding
Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-liteGoogle$0.25$1.501MResponding
Gemini 2.5 Flashgoogle/gemini-2.5-flashGoogle$0.30$2.501MResponding
Gemini 2.5 Progoogle/gemini-2.5-proGoogle$1.25$10.001MResponding
Gemini 3.5 Flashgoogle/gemini-3.5-flashGoogle$1.50$9.001MResponding
Gemini 3.1 Pro (preview)google/gemini-3.1-pro-previewGoogle$2.00$12.001MResponding
Gemini 3.6 Flashgoogle/gemini-3.6-flashGoogle1MResponding
GPT-5 Nanoopenai/gpt-5-nanoOpenAIResponding
GPT-5 Miniopenai/gpt-5-miniOpenAIResponding
GPT-5.4 Miniopenai/gpt-5.4-miniOpenAIResponding
GPT-5.6 Solopenai/gpt-5.6-solOpenAIResponding
GPT-5.6 Terraopenai/gpt-5.6-terraOpenAIResponding
GPT-5.6 Lunaopenai/gpt-5.6-lunaOpenAIResponding

How this is measured

Every hour we send an identical short prompt to each tracked model through a single gateway account and record what the call was billed. Dividing the billed input and output cost by the tokens consumed gives an observed price per million tokens.

Prices shown are therefore what one buyer was actually charged, at that hour, on that route. Enterprise agreements, batch discounts and cached-input rates will differ. Where a provider does not itemise cost, the price cell reads “—”.

A “not responding” model returned an error on the most recent sweep. That is a fact about our route to it, not necessarily a global outage.