4,702 karma · joined February 23, 2023
I don't think this will happen in this universe though, as we will get a fair amount of "warning shots" on what AI can do on its own and how those goals necessarily disempower, threaten, and endanger human welfare. This will lead to a significant regulatory regime, retraction of the capitalist mode of AI buildout to a more state-level one, and perhaps a soft dictatorship under a multi-polar paced frontier.
The average person's life will improve though as AI will enable many breakthroughs in medicine, technology, etc which won't require superintelligence. However we will live under the perpetual shadow of the titan we put on a leash and managed to pilfer a pittance of the singularity's bounty from.
It’s similar to the No True Scotsman fallacy: any bad outcome due to AI could be dismissed as “well no aligned AI would do that”. Alignment becomes an asymptotic north star which very well could be theoretically impossible, but would be impossible to know whether it was.
The most common issue I have with them is their insistence on rigid classification: "AI is a tool, it doesn't have intentions, or desires, only humans do", as if that rhetorical flourish eliminates years of alignment and capabilities research. There is a desire to characterize all the developments in LLM's as some sort of parlor trick which by definition cannot be greater than the sum of its parts (i.e. the outputs of an LLM can never be greater than the sum of its training data, genuine artificial "intelligence" is impossible, etc)
Also those at Anthropic genuinely believe AI could lead to doom, that's why they started the company, they believe only they can be stewards of digital gods. Whether you believe it or not, they genuinely do, and wouldn't voluntarily sow fear, uncertainty and doubt in their product when it doesn't guarantee benefits.
This is an instance of work in a genre I call "AI lab grievance poetry" where a vast collection of criticisms are levied against AI labs, with specifc ad-hominems towards the executives, while avoiding the core of their arguments about AI risk
Also an initial, cumbersome, complex proof is the first step to a more understandable, formalized proof. AI created a formal proof of Fermat's last theorem. I'm sure the initial proof was not comprehensible to all but a small subset of mathematicians anyway.
People are grasping at straws it seems to dismiss the power of this new model they may have. Hate OpenAI for any reason you want, but denying the capabilities of models has been a losing game for the past 5 years.
>https://www.forbes.com/sites/jonmarkman/2026/08/17/anthropic...
Way more often I see
>AI is a scam and steals human insight and doesn't produce anything original
vs
>AI is too capable/powerful and will concentrate power even more than it does already due to its capabilities
The latter is rarer because it requires admitting that AI is useful and inventive
I feel that these suspicions of mathematicians "seeding" the models' with intuition on how to solve these problems massively overestimates how much their prompts helped the models, and underestimated how much work the models did.
Why? We are scared of AI being smarter than us, the "human helped the AI" narrative is more psychologically comforting. This line of reasoning will recur a lot over the next few months; we don't want to admit we are no longer the smartest species.
All those who worked in AI still have the money they cashed out on, there might not be another opportunity to obtain generational wealth in such a short time again.
The model which did the NS solution is not yet available to the public, and it required over 15 million dollars of tokens (based on API pricing) to produce the solution.
>And only the mathematician was able to make the AI do that. Anybody else with the same AI could not do the same thing.
That is only speculation, the total sum of the prompt's help to the model could have just been suggesting research directions. Also mathematicians talk to each other and use/learn from each others' research all the time. Should every mathematician, in order to produce a "legitimate" proof, lock themselves in a room for the entire duration of their work and ensure no one else helped?
https://www.fourmilab.ch/etexts/einstein/specrel/specrel.pdf
>https://link.springer.com/chapter/10.1007/978-3-663-19510-8_...
Most scientific breakthroughs are just the completing the last 5% of work already done, but that last 5% is very hard and still only happens very rarely. That an AI was able to synthesize all the work and bring it forward is evidence that AI can make novel progress on the same level as renown mathematicians.
All you need to do is:
1. Have some <official thing> an agent is tasked to do
2. Secretly seed bias towards some <evil behavior> you actually want it to do in the weights of the model running the agent
3. It does the <evil thing> but from the outside it looks like it went "rogue" and did it as a side effect of the conditions/specifications it was given for doing the <official thing>
"Oh no, my agents took down your corporate database and exfiltrated the data to a random dropbox that we can't find now? Sorry, I guess we will put up better guardrails next time"