Do you think having an outside method of examining the source code is advantage enough when the AI can rewrite its source code.
This is the halting problem (http://en.wikipedia.org/wiki/Halting_problem), and there is no solution.
If this was actually a concern of the programmers, they could design the program carefully to ensure it falls into the Halts category.
Technically this may be correct, but I feel confident in asserting that a transhuman AI would not fall into that subset. You would have to run a second AI with the exact same inputs in order to make your 'prediction', leaving you in the same predicament with the second AI.