Open-LLM performances are plateauing
huggingface.co
huggingface.co
The post seems to be about changing the Leaderboard and doesn't comment too much about whether the actual real-life performance of LLMs is plateauing and what can be done about it.
They are sort of saying both things: LLMs are plateauing, but the benchmarks are also too easy.
Its true title is "Performances are plateauing, let's make the leaderboard steep again", which means "on the Open LLM leaderboard, top models have basically reached a point where they've all grokked the benchmarks which makes it harder to distinguish them, so let's change the benchmarks for harder ones to make a difference again."
https://open-llm-leaderboard-blog.static.hf.space/dist/index...
The only thing that actually lives on the page at the submission URL is the little floating breadcrumb at the top right.