← LLM 
OpenAI
gpt-5-mini
reasoning
Measured 2026-08-25
Score
71
Task done
89.12%
82–95% 95% CI bootstrap over items · n=95 · measured 2026-08-25
Refusal quality
83.3%
names the gap 83% · offers a route 83% · 10 probes x 3 iterations · 5 silent turns scored zero
Fabrication
0%
0 of 10 probes (30 probe runs) · ≤31%
Dead-air
67.1%
198 of 295 turns
Tool silence
100.0%
all of 19 items (198 tool-call turns) · ≥82%
Stalled
21.1%
19 of 90 runs
TTFT p50
746ms
746–794 p50–p90
Cost / 1M tok
$2.00
Share this result
Embed the live badge
Markdown
[](https://benchmarks.speko.ai/llm/openai-gpt-5-mini)
HTML
<a href="https://benchmarks.speko.ai/llm/openai-gpt-5-mini"><img src="https://benchmarks.speko.ai/badge/llm/openai-gpt-5-mini.svg" alt="Speko llm rank"></a>
URL
https://benchmarks.speko.ai/llm/openai-gpt-5-mini