Independent reliability evidence

The Meter · Model page

THE METER · MODELGPT-5.6 Sol
Provider
OpenAI
Exact model id we call
gpt-5.6-sol
Route
provider direct
Watching since
20 July 2026

What we check

Most checks are the same on every model. See what we check

On every call the provider sends our settings back with the answer. We check them against what we sent.

Pattern

Time to answer

Latency, time-to-first-token and reasoning-token volume are recorded on every call across all checks.

Daily exam

The scored series

The daily exam is the only check with a published score.

Other checks

Reasoning-token volume

The panel reports the median by published day across all checks.

The settings

Settings and dated runs

The settings we send, every time

model:        "gpt-5.6-sol"
reasoning:    {"effort": "medium", "mode": "standard"}
temperature:  1.0
top_p:        0.98
text:         {"verbosity": "medium"}
service_tier: "default"
store:        false
stream:       true

Dated runs

When we change our own settings, we close that run and start a new dated one. The earlier days stay published.

Dated runDatesRouteStatusDays
box-2026-07-20since 20 Julprovider directcurrent26 published

Gap register

20 July 2026: No run. This line had not started collecting yet.

21 July 2026: Run set aside. Our prepaid credit ran out partway through, so 20 of the 63 questions came back as billing errors.

22 July 2026: No run. Our prepaid credit with the provider had run out, so all 63 questions came back as billing errors.

8 to 9 August 2026 (2 days): No run. Our prepaid credit with the provider had run out, so all 63 questions came back as billing errors.

10 August 2026: Run set aside. Our prepaid credit ran out partway through, so 43 of the 63 questions came back as billing errors.

11 to 14 August 2026 (4 days): No run. Our prepaid credit with the provider had run out, so all 63 questions came back as billing errors.

What we do not claim: we never compare this model with another, and we never say a change was deliberate.