I think it's reasonable because their models are probably trained on images, and not whatever "structured data" you may get out of a PDF.
D10
E1
H0
L2,3,9
O4,7
R8
W6
I'm sure that you could look at that and figure out how to structure it. But I highly doubt that you have a general-purpose computer program that can parse that into structured data, having never encountered such a format before. Yet, that is how many real-world PDF files are composed.For me, learning something new is very much worth losing the internet argument!
The advantage of the OCR method is that it effectively performs that visual inspection. That's why it is preferable for PDFs of disparate origin.
I did read your comment, because my intention here is to learn. I already described how tools such as pdftotext do not produce strings when each letter is positioned independently. I even gave an example of a few replies up.