The Meter · Model page
- Provider
- OpenAI
- Exact model id we call
- gpt-5.6-sol
- Route
- provider direct
- Watching since
- 20 July 2026
What we check
Most checks are the same on every model. See what we check ↗
On every call the provider sends our settings back with the answer. We check them against what we sent.
Pattern
Time to answer
Latency, time-to-first-token and reasoning-token volume are recorded on every call across all checks.
Daily exam
The scored series
The daily exam is the only check with a published score.
Other checks
Reasoning-token volume
The panel reports the median by published day across all checks.
The settings
Settings and dated runs
The settings we send, every time
model: "gpt-5.6-sol"
reasoning: {"effort": "medium", "mode": "standard"}
temperature: 1.0
top_p: 0.98
text: {"verbosity": "medium"}
service_tier: "default"
store: false
stream: trueDated runs
When we change our own settings, we close that run and start a new dated one. The earlier days stay published.
| Dated run | Dates | Route | Status | Days |
|---|---|---|---|---|
| box-2026-07-20 | since 20 Jul | provider direct | current | 26 published |
Gap register
20 July 2026: No run. This line had not started collecting yet.
21 July 2026: Run set aside. Our prepaid credit ran out partway through, so 20 of the 63 questions came back as billing errors.
22 July 2026: No run. Our prepaid credit with the provider had run out, so all 63 questions came back as billing errors.
8 to 9 August 2026 (2 days): No run. Our prepaid credit with the provider had run out, so all 63 questions came back as billing errors.
10 August 2026: Run set aside. Our prepaid credit ran out partway through, so 43 of the 63 questions came back as billing errors.
11 to 14 August 2026 (4 days): No run. Our prepaid credit with the provider had run out, so all 63 questions came back as billing errors.
What we do not claim: we never compare this model with another, and we never say a change was deliberate.