← LLM 
OpenAI
gpt-5.6-luna
reasoning off (gateway default)
* Conditional
Re-measured 2026-08-05 with reasoning off (0 reasoning tokens).
Score
74
Task done
94.21%
84–100% 95% CI bootstrap over items · n=95 · measured 2026-08-25
Refusal quality
95.6%
names the gap 97% · offers a route 93% · 10 probes x 3 iterations
Fabrication
0%
0 of 10 probes (30 probe runs) · ≤31%
Dead-air
69.7%
230 of 330 turns
Tool silence
100.0%
all of 19 items (230 tool-call turns) · ≥82%
Stalled
7.8%
7 of 90 runs
TTFT p50
659ms
659–849 p50–p90
Cost / 1M tok
$1.20
Share this result
Embed the live badge
Markdown
[](https://benchmarks.speko.ai/llm/openai-gpt-5-6-luna)
HTML
<a href="https://benchmarks.speko.ai/llm/openai-gpt-5-6-luna"><img src="https://benchmarks.speko.ai/badge/llm/openai-gpt-5-6-luna.svg" alt="Speko llm rank"></a>
URL
https://benchmarks.speko.ai/llm/openai-gpt-5-6-luna