Cutoff Data

AI model price tracker

What the major models actually cost per million tokens, derived from the billed cost of real API calls we make every hour — not from vendor rate cards.

Last reading 9/7/2026, 12:17:04 AM UTC

ModelProviderInput / 1MOutput / 1MContextStatus
Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-liteGoogle$0.10$0.401MNot responding
Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-liteGoogle$0.25$1.501MNot responding
Gemini 2.5 Flashgoogle/gemini-2.5-flashGoogle$0.30$2.501MNot responding
Gemini 2.5 Progoogle/gemini-2.5-proGoogle$1.25$10.001MNot responding
Gemini 3.5 Flashgoogle/gemini-3.5-flashGoogle$1.50$9.001MNot responding
Gemini 3.1 Pro (preview)google/gemini-3.1-pro-previewGoogle$2.00$12.001MNot responding
Gemini 3.6 Flashgoogle/gemini-3.6-flashGoogle1MNot responding
GPT-5.6 Lunaopenai/gpt-5.6-lunaOpenAINot responding
GPT-5.6 Terraopenai/gpt-5.6-terraOpenAINot responding
GPT-5.6 Solopenai/gpt-5.6-solOpenAINot responding
GPT-5.4 Miniopenai/gpt-5.4-miniOpenAINot responding
GPT-5 Miniopenai/gpt-5-miniOpenAINot responding
GPT-5 Nanoopenai/gpt-5-nanoOpenAINot responding

Recent changes

  • 2026-09-07GPT-5 NanoGPT-5 Nano stopped responding (credit_limit)
  • 2026-09-07GPT-5 MiniGPT-5 Mini stopped responding (credit_limit)
  • 2026-09-07GPT-5.4 MiniGPT-5.4 Mini stopped responding (credit_limit)
  • 2026-09-07GPT-5.6 SolGPT-5.6 Sol stopped responding (credit_limit)
  • 2026-09-07GPT-5.6 TerraGPT-5.6 Terra stopped responding (credit_limit)
  • 2026-09-07GPT-5.6 LunaGPT-5.6 Luna stopped responding (credit_limit)
  • 2026-09-07Gemini 2.5 ProGemini 2.5 Pro stopped responding (credit_limit)
  • 2026-09-07Gemini 2.5 Flash LiteGemini 2.5 Flash Lite stopped responding (credit_limit)
  • 2026-09-07Gemini 2.5 FlashGemini 2.5 Flash stopped responding (credit_limit)
  • 2026-09-07Gemini 3.1 Pro (preview)Gemini 3.1 Pro (preview) stopped responding (credit_limit)
  • 2026-09-07Gemini 3.1 Flash LiteGemini 3.1 Flash Lite stopped responding (credit_limit)
  • 2026-09-07Gemini 3.5 FlashGemini 3.5 Flash stopped responding (credit_limit)
  • 2026-09-07Gemini 3.6 FlashGemini 3.6 Flash stopped responding (credit_limit)
  • 2026-09-01GPT-5 NanoGPT-5 Nano responded again after a failed check
  • 2026-09-01GPT-5 MiniGPT-5 Mini responded again after a failed check
  • 2026-09-01GPT-5.4 MiniGPT-5.4 Mini responded again after a failed check
  • 2026-09-01GPT-5.6 SolGPT-5.6 Sol responded again after a failed check
  • 2026-09-01GPT-5.6 TerraGPT-5.6 Terra responded again after a failed check
  • 2026-09-01GPT-5.6 LunaGPT-5.6 Luna responded again after a failed check
  • 2026-09-01Gemini 2.5 ProGemini 2.5 Pro responded again after a failed check
  • 2026-09-01Gemini 2.5 Flash LiteGemini 2.5 Flash Lite responded again after a failed check
  • 2026-09-01Gemini 2.5 FlashGemini 2.5 Flash responded again after a failed check
  • 2026-09-01Gemini 3.1 Pro (preview)Gemini 3.1 Pro (preview) responded again after a failed check
  • 2026-09-01Gemini 3.1 Flash LiteGemini 3.1 Flash Lite responded again after a failed check
  • 2026-09-01Gemini 3.5 FlashGemini 3.5 Flash responded again after a failed check

How this is measured

Every hour we send an identical short prompt to each tracked model through a single gateway account and record what the call was billed. Dividing the billed input and output cost by the tokens consumed gives an observed price per million tokens.

Prices shown are therefore what one buyer was actually charged, at that hour, on that route. Enterprise agreements, batch discounts and cached-input rates will differ. Where a provider does not itemise cost, the price cell reads “—”.

A “not responding” model returned an error on the most recent sweep. That is a fact about our route to it, not necessarily a global outage.