Some customers will tell the AI to do very bad things (substitute here whatever you consider very bad). Should it do those things?
One possible answer: yes, it should do all the very bad things, and we will take care of prosecuting the customer later, once the bad thing is accomplished. Is this your answer?