> Luna without reasoning turned on might as well be a model from early 2025. This is not how anyone is using this.
Importantly, it's probably also not what it's been trained to do.
Until quite recently, OpenAI used to ship dedicated "instant" and "reasoning" models. Newer ones seem to have reasoning levers that can be turned down all the way to zero, but that doesn't mean they don't take a significant performance hit when doing that.