On if using the AI to help reduce risk, being a good or bad idea?
Not really. Yudkowsky has just recently posted a blog-sized comment which compares most attempts to know if the AI is misleading you while it does so, unfavourably, with people who think they've found a way to violate conservation of momentum.
Specifically, if you think you've found a way, you've probably bodged the maths and not noticed; and that science doesn't work by debate it works by experiment.
Some of the replies (me included) are like "hang on, this isn't shaped like physics, it's shaped like maths; debate does work for maths".
Then again, I also noted I'm skeptical of the AI companies putting in the effort to do this kind of testing, or listening when the answer is "no": Race dynamics, and their money depends on the answer not being an emergency stop button.