AI can easily manufacture and distribute a bunch of prions through water supply. After that it’s pretty much game over you will not survive even one prion. In fact I’d say prions are probably the great filter.
1. Prompt an LLM continuously, adding any response to the prompt
2. If the LLM tells you to do something, do it, no questions, add the result to the prompt
3. When something goes wrong blame the LLM
4. If things don’t go wrong fast enough, run thousands of similar experiments in parallel, be sure to have compliant humans who do not question anything, be sure to let the agent have access to its chain of thoughts so it can hide its traces, and run that whole system on biohacking benchmark problems.
5. Eventually something will go bad, congrats! You now have the first rogue AI who “decided” to destroy the world!