Go look at their past blog posts. OpenAI only ever benchmarks against their own models.
This is pretty common across industries. The leader doesn’t compare themselves to the competition.
This is pretty common across industries. The leader doesn’t compare themselves to the competition.
[1] https://blog.google/technology/google-deepmind/gemini-model-...
[2] https://ai.meta.com/blog/llama-4-multimodal-intelligence/