57 karma · joined May 2, 2023
@nipple_nip on twitter
GLM-4.7 like a mix of Sonnet 4.5 and GPT-5 (the first version not the later ones). It has deep deep knowledge, but it's often just not as good in execution.
They're very cheap to try out, so you should see how your mileage varies.
Ofcourse for the hardest possible tasks that GPT 5.2 only approaches, they're not up to scratch. And for the hard-ish tasks in C++ for example that Opus 4.5 tackles Minimax feels closer, but just doesn't "grok" the problem space good enough.
[0] - https://en.wikipedia.org/wiki/Workers%27_self-management#Yug...
It is ultimately all speculation, until Deepseek releases their own 145B MoE model, and then we can compare the activations/results
And I also can't wait to see how much Phind will improve further if the Glaive dataset is added onto it.
Edit: Contrastive search, dynamic temperatures.
There's currently multiple attempts at creating what you describe as books4.
All in all, this is either a honeypot, or somebody trying to discredit Kosovo.
Still, for certain languages, only libgen and public piracy websites contain any scientific or fiction material in digital formats. E.g. my native language doesn't have easily accessible e-books at all, unless you go through illegal means.
I hope somebody undertakes the steps necessary to train on the entirety of libgen. The amount of high quality tokens in libgen should be substantial.