If one creates a human-level AGI with certain human-friendly goals, and allows it to self-modify freely, the odds are high that it will eventually self-modify into a condition where it no longer pursues the same goals it started out with
I think their idea is a rational agent would have no incentive to change its terminal goals, as changing your utility function has to be of negative utility save for in some extreme edge cases that aren't likely to be relevant. The hard part is giving it the correct terminal goal.
I also don't really think Bostrom is anthropomorphising. Basing your AI ideas on Piaget's theory of cognitive development, that seems like anthropomorphism - though likely entirely appropriate when the human mind is what you're attempting to mimic.
Interesting article, brings back fond memories of Sl4,reading Ben, Yudkowsky, Robin Hanson, Gwern, and two Satoshi candidates argue was quite fun. I was convinced then that by 2015 nanotech or AI would have eaten the world. I'm a little more uncertain now, but we still have a couple days.