I legitimately dont know how to reply, bc by this point llms co-own all aspects of my life and jumps between gpt4->claude3->claude3.5->o1 have all been very noticeable
23 karma · joined January 11, 2012
it makes sense to benchmark them infependently since prompt processing is done in parralel for each token and is compute bound and token generation is sequential and bound by memory banwidth
Now multiply this by several billions.
Should be glorious
And you are just making unjustified bold statements
Well, so was Ukraine, and look at it now.
And your last point is wrong. ML models are studied and understood much better than human reasoning.