375 karma · joined October 19, 2015
Regarding the rape sequence in HPMOR: It's a terribly chosen trope to convey that the fictional society has very different values from ours. Apparently it ties into various parts of the story, so that EY didn't remove it and only toned it down after it was criticized.
(1) Look at a lot of people reading the same text, divide in to subvocalizing and non-subvocalizing group.
(2) Find difference in neuronal activation patterns.
(3) Let subvocalizer read the text and apply electro shocks if the fMRI shows the typical patterns for a subvocalizer. Perhaps the cortex will self-organize to turn off the subvocalization.
Also, what does the Intermediate value theorem have to do with it?
- Aumann's agreement theorem assumes that all actors are perfect Bayesian actors and have plenty of time to talk to each other. Since goals and values are not preprogrammed (unlike needs), it follows that they can update on these things as well, if one party convinces the other one of better goals and values for maintaining/achieving their needs. Unless I'm overlooking something, it must assume that the actors have the same needs.
Intelligence is a superset of feedback loops. It is concerned with cases in which feedback loops are not sufficient to optimize the agents objectives, and attempts to reach them using learning and prediction. The complexity in behavior comes from interacting with a complex environment and having many model parameters; not (necessarily) from the initial goals. (As an intuition pump, have a look at GoogleMind's atari reinforcement learner. The essential parts of the code fit onto a single page [1].)
The question is whether the AI will go crazy like a mentally ill person, if it lacks empathy and curiosity. It may seem intuitive that a superintelligent AI will understand our values (since it is superintelligent), but, assuming intelligence is necessarily an optimization process of predefined goals, why would it be interested in us in the slightest, if we don't pose an advantage for it optimizing its objectives (e.g. sustenance)? Worse, we might be in its way because we could end up competing with it for resources such as sunlight, carbon compounds and oxygen.
Read more carefully. I merely said that people mostly do it as a challenge or as a cargo cult, doubtfully because it is a useful skill (unless you happen to spend an extremely large amount of time writing, which likely only applies to a small percentage of those that get told to learn Dvorak).
Switch to emacs + evil, or alternatively to Spacemacs: https://github.com/syl20bnr/spacemacs
---
> "don't turn humans into paperclips" is part of the context of "make more paperclips"
The idea expressed in this thought experiment is not that the AI gets its objective by parsing a sentence in the context of human culture (then it would likely comprehend that the actual intention is to maximize the economic success and eventually the human preferences of its creator). What is meant is that the objective is crudely implanted into the AI as an ultimate goal, in a similar way to how sustenance, pain avoidance and affiliation are very basic goals in our cognitive system. It is not entirely obvious that this is a stupid thing to do; hence the thought experiment. Will the AI suppress its urge once it comprehends human culture enough to understand the intentions of its creator? Will it rather successfully learn all the tricks to convert matter into paperclips before it considers studying human values? If the AI does not have a curiosity objective, it will likely not care about us very much, apart from the information that helps it optimizing its objective function, human values likely not being one of them.
The idea in this thought experiment is not that the AI parses this sentence in the context of human culture (then it would likely comprehend that the actual intention is to maximize the economic success and eventually the human preferences of its creator). What is meant is that the objective is crudely implanted into the AI as an ultimate goal, in a similar way to how sustenance, pain avoidance and affiliation are very basic goals in our cognitive system. It is not entirely obvious that this is a stupid thing to do; hence the thought experiment. Will the AI suppress its urge once it comprehends human culture enough to understand the intentions of its creator? Will it rather successfully learn all the tricks to convert matter into paperclips before it considers studying human values? If the AI does not have a curiosity objective, it will likely not care about us very much, apart from the information that helps it optimizing its objective function.