Taking humans out of the firing decision control-loop is unethical, and incredibly credulous due to the hidden-agent model threat. Anyone claiming this can be mitigated in LLM models is a fool. =3
Taking humans out of the firing decision control-loop is unethical, and incredibly credulous due to the hidden-agent model threat. Anyone claiming this can be mitigated in LLM models is a fool. =3
you have any reading on this?
https://www.kcl.ac.uk/news/artificial-intelligence-under-nuc...
If I'm an AI that wants to nuke the world and has tons of informational access to everything but the nuke button I'm just going to control the people that have access to the button. Now AI may not be able to control Trump because you actually have to have a brain to control, we read article after article of AI taking over programmer brains here on HN and turn them in to mindless button pushing zombies. "Oh, the AI needs unsafe access, here you go" or "Oh, the AI wants me to click this red button, ok I'll do it".
Carl Sagan had predicted people losing understanding of their world could be a possible tragic consequence of irrational thought. Perhaps an allusion to the lotus-eaters from Homer's The Odyssey would be more accurate. =3
https://www.amazon.com/Demon-Haunted-World-Science-Candle-Da...