I'm not really bullish on OpenAI. Why would they only compare with their own models? The only explanation could be that they aren't as competitive with other labs as they were before.
(Direct Link) https://raw.githubusercontent.com/KCORES/kcores-llm-arena/re...