Turns out he was just writing LLM prompts way ahead of his time.
Turns out he was just writing LLM prompts way ahead of his time.
I don’t think that’s the main problem, there are a lot of moral dilemmas where even humans can’t agree what’s right.
If each human could pause the state of the world and gather all information and then decide, they would act humanely
Just think of abortion, wars or legalizing drugs. People disagree completely over those because nobody agrees which choice would be the moral one.
for some reason that one wasn't even included in the list of books in the series on the inside jacket of the other books that I had.
I remember I had to really hunt for it and it was from a different publisher. never knew why.
HPMOR offers a solution called 'coherent extrapolated volition' – ordering the super intelligent machine to not obey the stated rules to the letter, but to act in the spirit of the rules instead. Figure out what the authors of the rules would have wished for, even though they failed to put it in writing.
We are debating scifi, of course.
What if the original author was from long ago and doesn't share modern sensibilities? Of course you can compensate when formulating them to some extent, but I imagine there will always be potential issues.
I recall that story of the guy who tried to use AI to recreate his dead friend as a microwave and it tried to kill him[0].
You couldn't sell a sci-fi story where AIs just randomly go insane sometimes and everyone just accepts it as a cost of doing business, and because "humans are worse," but that's essentially reality. At least not as anything but a dark satire that people would accuse of being a bit much.
[0]https://thenextweb.com/news/ai-ressurects-imaginary-friend-a...
Certainly AI safety isn't perfect, but if you're going to criticize it at least criticize the AIs people actually use today. It's like arguing cars are unsafe and pointing to an old model without seatbelts.
It's not surprising at all that people are willing to use AIs even if they give dangerous answers sometimes, because they are useful. Surely they're less dangerous than cars or power tools or guns, and all of those have legitimate uses which make them worth the risk (depending on your risk tolerance.)
And now I wouldn’t even trust them to understand the laws 100% of the time.
Or so says Ted Chiang: https://en.m.wikipedia.org/wiki/The_Lifecycle_of_Software_Ob...
Like seemingly all torment nexii, the warning part of the tale is forgotten.