GPT-4 is clearly very impressive, they should show it off honestly and transparently. Instead OpenAI clearly treats these evaluations as a part of their sales and marketing, with inflated claims to match.
A relatively tiny percentage of people have read this relatively obscure academic paper, an even smaller amount care, and probably none will stop using ChatGPT or investing in OpenAI because of it.
Meanwhile its performance on the bar exam made international headlines which millions if not bilions of people have read.
A lie is halfway across the world before the truth gets out of bed.