Intuitively, this makes some sense - one would expect an object classifier to not care too much about determining viewpoint, so the amplified representation of a dog or a slug is flat. I think convolution layers being the bottom-most layer also has something to do with it.
My dreams are a lot more perspective-correct though. Deepdream certainly entertains the idea that biological dreaming might be somehow similar to gradient ascent. Even if it were so, it means that the sensory experiences we feel in our dreams somehow integrate a much more unified "reality" than what we would experience if we were only dreaming with an object classifier.