Feedforward CNNs cannot tolerate as much noise and real-world variability as RNNs.
Recurrence can help with robustness in some other very important ways as well.
Citations for this dates from the 80s and 90s. I don't know the best reference offhand. You could look at some old Hinton stuff if you're a fan. Lots published on this.
Nothing like this has been published AFAIK.
After you have the results of this experiment you can try to explain them with attractors and what not, but I would be surprised if there was much difference. Would make a good paper though!
Unless you're dealing with opinions, I disagree. The onus is on the person trying to give evidence to actually give evidence.