Problem is, this paper won't be able to produce those results from historical photos. I find them quite misleading, to be honest.
You'll need to either use a separate 3D depth estimation AI, or (more likely) have someone do a manual stereoscopic 3D conversion of your historical image. Only then (when you have depth data) can the algorithm presented in this paper start its work.