A new paper from OpenAI (Sept 2025) makes a compelling argument that the stubborn problem of LLM hallucination isn't a mysterious glitch or something that can be solved with more scale alone. I wrote a deeper analysis of this idea and what it means for the future of AI evaluation .