Oh, another LLM skepticism paper from Apple.
This paper from last year doesn't age well due to rapid proliferation of reasoning models.
This paper from last year doesn't age well due to rapid proliferation of reasoning models.
While everyone learned the bitter lesson, apple chose to focus on small on-device models even after the explosion of chatgpt.
"See this is why we can't build with transformers and had to use JEPA and look how much better it is!"