Some prior discussion prompted by "Why LLMs Suck at OCR": https://news.ycombinator.com/item?id=42966958
Even with extremely blurry typewriter scans that are difficult for me to decipher.
It's incredible.
I'm sure there's cases where it will fail but just OCRing 90% of the files would be a big win.