Extracting Structured Data from Templatic Documents
ai.googleblog.com
ai.googleblog.com
Mind you this was a few years ago and I was primarily testing with pytesseract. I would be curious if this team actually used the Google OCR API or an internally tuned one that isn't GA, and how that differs FOSS Tesseract.