When people prompt these AIs they tend to steer them towards what they want to see. A concerned researcher prompts it towards concerning content. A starry eyed one will avoid anything close to problematic. The interesting, and actually problematic section lies in the middle, much like you describe here. What would the AI do when asked to defend Alex Jones and his lies against the parents of the children dead at sandy hook?
The real terror of AI is what it allows humans to do to each other, not what it will do by itself.