← LLM

OpenAI

gpt-5.6-luna

reasoning off (gateway default)

* Conditional

Re-measured 2026-08-05 with reasoning off (0 reasoning tokens).

Score

74

Task done

94.21%

84–100% 95% CI bootstrap over items · n=95 · measured 2026-08-25

Refusal quality

95.6%

names the gap 97% · offers a route 93% · 10 probes x 3 iterations

Fabrication

0%

0 of 10 probes (30 probe runs) · ≤31%

Dead-air

69.7%

230 of 330 turns

Tool silence

100.0%

all of 19 items (230 tool-call turns) · ≥82%

Stalled

7.8%

7 of 90 runs

TTFT p50

659ms

659–849 p50–p90

Cost / 1M tok

$1.20

Share this result

Embed the live badge

Speko llm rank
Markdown
[![Speko llm rank](https://benchmarks.speko.ai/badge/llm/openai-gpt-5-6-luna.svg)](https://benchmarks.speko.ai/llm/openai-gpt-5-6-luna)
HTML
<a href="https://benchmarks.speko.ai/llm/openai-gpt-5-6-luna"><img src="https://benchmarks.speko.ai/badge/llm/openai-gpt-5-6-luna.svg" alt="Speko llm rank"></a>
URL
https://benchmarks.speko.ai/llm/openai-gpt-5-6-luna