Cutoff Data
Model speed & uptime index
How fast each major model answers, and how often it fails, measured hourly from an independent client over the last 7 days.
Last reading 8/7/2026, 6:17:31 PM UTC
| Model | Provider | Median | 95th pct | Uptime | Probes | Now |
|---|---|---|---|---|---|---|
| Gemini 2.5 Flash Litegoogle/gemini-2.5-flash-lite | 846ms | 853ms | 100% | 2 | OK | |
| GPT-5.6 Lunaopenai/gpt-5.6-luna | OpenAI | 1.06s | 1.25s | 100% | 2 | OK |
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1.09s | 1.41s | 100% | 2 | OK | |
| GPT-5.6 Solopenai/gpt-5.6-sol | OpenAI | 1.35s | 1.60s | 100% | 2 | OK |
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.39s | 1.44s | 100% | 2 | OK | |
| GPT-5.4 Miniopenai/gpt-5.4-mini | OpenAI | 1.47s | 2.08s | 100% | 2 | OK |
| GPT-5.6 Terraopenai/gpt-5.6-terra | OpenAI | 1.84s | 2.30s | 100% | 2 | OK |
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 2.13s | 2.43s | 100% | 2 | OK | |
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 2.57s | 2.87s | 100% | 2 | OK | |
| GPT-5 Nanoopenai/gpt-5-nano | OpenAI | 2.81s | 3.15s | 100% | 2 | OK |
| GPT-5 Miniopenai/gpt-5-mini | OpenAI | 3.05s | 3.33s | 100% | 2 | OK |
| Gemini 2.5 Progoogle/gemini-2.5-pro | 4.00s | 4.21s | 100% | 2 | OK | |
| Gemini 3.1 Pro (preview)google/gemini-3.1-pro-preview | 4.08s | 4.17s | 100% | 2 | OK |
How this is measured
Once an hour we send every tracked model the same short, deterministic prompt and record wall-clock time from request to completed response, plus whether the call succeeded.
Median and 95th-percentile figures are computed across all successful probes in the trailing 7-day window. Uptime is successful probes divided by total probes.
These are single-client measurements over one network route, at short prompt lengths. They are a fair basis for comparing models against each other under identical conditions, not a substitute for a provider’s own status page.