529 karma · joined January 19, 2019
Bravo to the author. This is inspiring work.
A bash script can clone and build stockfish, feed in human moves, and reply. By your standard, this bash script would "destroy any human at chess."
Are you interested in assessing the intelligence of the model, or the intelligence of the tools the model can use?
https://arxiv.org/html/2509.24239v4
Researchers asked frontier models to play chess. Have a look at the MAR rates in Table 3. When not explicitly told which moves were legal, no model identified legal moves at a rate better than 80%. Many asked for more illegal moves than legal moves. And even when explicitly told which moves were legal, the models continued to ask for illegal moves. With illegal asks discarded, none of the bots could beat a chess model calibrated to 1100 ELO.
The author of the originating post says that "current frontier models need laborious oversight and guardrails on even the simplest tasks", and he's absolutely correct.
Technically correct. Losing to random play is definitely not master level.
I don't care if it's interesting or not. A bot too stupid to play chess is too stupid to be a threat to me. So I reject Dario's fearmongering and that of anyone allied with him.
Disconnect it from the internet, where it cannot copy from humans, and no, it certainly will not. If it can copy code from humans, then maybe, with guidance. But if a human is not there to offer guidance, no, I don't think it will improve on stockfish.
Worse, if it cannot copy code from humans, it will fail embarrassingly at playing chess on its own. This April 2026 paper found that frontier models which did not have specialized (human guided) tailoring could not beat an amateur-level bot (1100 ELO), and some were so bad they lost to random play.
https://arxiv.org/html/2509.24239v4
Look at the MAR rates in table 3. When not explicitly provided with a choice of legal moves, no bot could ask for legal moves at a rate better than 80%. Many of them asked for more illegal moves than legal moves. All of them continued to request illegal moves even when explicitly asked to choose from legal moves. That's not just less than perfect. That's abysmal.
So, recapping, if the most advanced bots on the planet, in April 2026, are not allowed to copy code written by humans, and are forced to play chess by themselves, they will embarrass themselves. And you and others are asking me to believe they are a threat to me somehow. That kind of talk is disconnected from reality.
What is common with all of the "impressive" examples you gave me? The bot had material to copy from, or a human standing nearby to tell it when it screwed up and suggest a path forward. They are not intelligent. They're so stupid, in fact, I could not trust them to balance my checkbook without knowing guard rails are in place. So spare me the hysterics.
"When do you draw the line of LLM is just some stupid thing vs. AI/AGI?" I'm not sure, but I do know the line is well beyond playing a game of chess without requesting illegal moves 20% of the time. There are 5-year-olds who can do that.
You need to ask yourself some hard questions.
1) Why have I been conditioned to expect that given enough compute, a bot can build anything, unguided?
2) Why have the embarrassing details about bot performance been hidden from me, or at least drowned out by hysterical claims about their wild successes? Why is the human role in these successes never highlighted?
3) What do the people training these bots stand to gain if they convince 51% of the voting population that bots are actually a meaningful threat?
Well, we agree that the results are piss poor, at least.
LLMs cannot do this. They still play very poorly relative to human pros, and even modern LLMs still understand the rules so poorly that they request illegal moves.
Anyone claiming that we should be afraid of a text generator that can't understand the rules of chess is being ridiculous.
Always claims and promises about what they "can" do. This pattern is constant across AI hype. No one can prove that an LLM will forever fail to solve problem X, so the claim goes uncontested, and humans are biased to believe unrebutted claims.
I reject this fallacious line of reasoning. I will instead require evidence of what they "have" done. No LLM to date has written a chess engine that beats stockfish. More importantly, no LLM has ever written any chess engine without heavily copying code and patterns from humans.
Watt for watt, every one of these models loses to stockfish. There is no possible scenario where we need to be afraid of "rogue superintelligence" when the intelligence in question cannot reason or plan well enough to play chess.
Lecun has the correct approach. Mock these people relentlessly for their attention-seeking doomerism.
Unsubsidized is the more accurate word here. Some governments have chosen not to pay public money to stock these books in libraries, but no government has created criminal penalties for ownership.
It may be the case that her library includes some books that genuinely carry criminal penalties, but the article does not provide enough info to assess that.
I tried to use claude for a completely menial task yesterday, adding some resistant-starch foods to my diet, and it was worse than useless. Actively wasting my time with claims that buckled as soon as I questioned them and completely falsified citations from product pages that it recommended.
If you are branding your product as "AI" when these word salad spitters are still dominating headlines, you deserve all the headwinds coming to you.
So I decline to adopt your convention of classifying it as an assumption. It might be more accurate to call it an observation.
I am inexplicably and disproportionately irritated by people like this. They've an a priori commitment to epistemic relativism. Unnaturally allergic to any claim that certainty is possible, and dedicated to undoing the careful work of those who are building up what may be known.
No, not every statement requires assumptions in order to hold. "The assumptions A implies B and B implies C, taken together, yield A implies C." This statement contains assumptions and makes observations about them, but it is true regardless of whether the assumptions it describes are true. The statement as a whole is "true" in the exact sense that the no counterexample to it can ever be given in any universe, under any set of assumptions.
I cannot wait for the day that Lean and other proof systems become accessible enough, conversant enough, and interdisciplinary enough to put these eternal confusion peddlers out of business.
I've found a lot to enjoy about programming with Rust. Some minor complaints:
- The syntax does not always make clear whether copy or move will happen on assignment. You need to manually go look at the object signature to see if it implements copy. This bit me once or twice.
- std::time::Duration as_millis() returns a u128, but from_millis() requires a u64, so you have to cast.
- Rust object methods accept a &Self param, but this is hidden when you call it. This convention exists in other languages, so I'm not sure where it originated, but it is a minor irk for me that the function signature and call don't match.
For example, leave the existing prefix binding (ctrl-b), but also add something nicer for day-to-day use (ctrl-space or similar).