Are you suggesting that detecting text from these mediocre-at-best photos is easy or that it's easy to determine which of the two sides is the control photo?
I would even go as far to argue that if these become widely used, we're going to see algorithmic "solvers" for this captcha in a matter of weeks.
The house numbers are easier for a computer to classify than the messy, weird contorted letters.