Anthropic needs some model with a fancy name so they can pretend for another while that their model is so powerful it will destroy the world if released. I propose Claude Legend 6.
Anthropic needs some model with a fancy name so they can pretend for another while that their model is so powerful it will destroy the world if released. I propose Claude Legend 6.
With how unsustainable they are they really need to hype up those bagholders.
This is not just for security, too. Using Fable for anything that could be remotely construed as being connect to chemistry or biology was impossible until a few weeks ago. Now it is slightly better, but still fails on many completely innocuous projects.
So as other models advance, Anthropic's sole frontier offering to entire academic fields remains an Opus that seems to get worse in capability each release. They're starting to become a joke in my field: at a conference a few weeks ago, one presenter laughed when I asked about his use of Fable and pointed out that it would downgrade if the letters 'd', 'n', and 'a' were anywhere near each other, which is not that far from my experience.
Is it, though?
We don't know what Mythos is really capable of, beyond what Anthropic has told us, and some second-hand accounts from orgs that have been whitelisted.
What we do know is that their withholding it from the masses is causing a lot of harm to their reputation and general annoyance. And probably a lot of money as well, as those people cancel their subscriptions in favor of other models. They are about to IPO, and you don't want people to have a bad taste in their mouth during this critical period.
As such, I think it is reasonable conclude that there must in fact be very valid reasons for them to keep going down this path of gradual access-widening. I'm never going to blindly trust a corporation, but in this case I'm not going to hate on them either because, at the risk of repeating myself, we just don't have all the facts.
Yes, it truly is.
Open models are extremely capable, as benchmark after benchmark has indicated.
Beyond that, for the vast majority of software development (including cybersecurity), the open models are there already. All that without having to pay the hefty Anthropic premium, not to mention all their bullshit with pretending their model is some sort of WMD and their awful uptime (although, to their credit, they seem more stable than Github).
I cannot fathom why anyone uses their service.
At work I am stuck with Claude because it is what the employer provides.
In my home setup I use a combination of different models; mostly GLM-5.3, DS-V4-Flash, and MiMo-2.5-pro.
I vastly prefer my home setup.
You asked if that is true, and speculated about mythos.
Uninterested in speculating about mythos, I explained why it is irrelevant anyway.
This was a successul conversation.