← LLM

OpenAI

gpt-5-mini

reasoning

Measured 2026-08-25

Score

71

Task done

89.12%

82–95% 95% CI bootstrap over items · n=95 · measured 2026-08-25

Refusal quality

83.3%

names the gap 83% · offers a route 83% · 10 probes x 3 iterations · 5 silent turns scored zero

Fabrication

0%

0 of 10 probes (30 probe runs) · ≤31%

Dead-air

67.1%

198 of 295 turns

Tool silence

100.0%

all of 19 items (198 tool-call turns) · ≥82%

Stalled

21.1%

19 of 90 runs

TTFT p50

746ms

746–794 p50–p90

Cost / 1M tok

$2.00

Share this result

Embed the live badge

Speko llm rank
Markdown
[![Speko llm rank](https://benchmarks.speko.ai/badge/llm/openai-gpt-5-mini.svg)](https://benchmarks.speko.ai/llm/openai-gpt-5-mini)
HTML
<a href="https://benchmarks.speko.ai/llm/openai-gpt-5-mini"><img src="https://benchmarks.speko.ai/badge/llm/openai-gpt-5-mini.svg" alt="Speko llm rank"></a>
URL
https://benchmarks.speko.ai/llm/openai-gpt-5-mini