Paper 0 · Methodology
Result confirmed / narrowA token budget censors a measurement by the cost of the correct answer.
Cap how much a model is allowed to write and it looks worse, without having changed at all. Our first study measured that effect on four OpenAI endpoints.
Run date 10 Jul 2026Four OpenAI endpoints tested
Read the paper ↗