Cutoff Data
GPT-5 Mini: price, speed and score
Everything we have measured about GPT-5 Mini from OpenAI, taken from real billed API calls rather than a published rate card.
Last reading 8/7/2026, 5:06:15 PM UTC
Input / 1M
—
Output / 1M
—
Context
—
Median response
3.71s
p95 response
4.37s
Uptime, 7 days
100%
Benchmark score
100%
Cost per task
—
Checks, 7 days
4
Compare with
How this is measured
Every hour we send one identical prompt to GPT-5 Mini through a single gateway account, record how long the call took and what it was billed, then divide that cost by the tokens consumed to get an observed price per million tokens.
Benchmark figures come from a fixed suite of code-graded tasks re-run against every tracked model. Scores are the share of available points earned; cost is the average spend per task.
These are one buyer's readings on one route. Enterprise pricing, batch discounts and cached-input rates will differ.