Hey, Tabula maintainer here. tabula-java only works with "vector" PDFs. That is, tables drawn with vector lines, squiggles and glyphs.
Integrating an OCR library is something we always wanted to do.
Integrating an OCR library is something we always wanted to do.
Here is a gif of table detection for a scanned PDF doc (the first run is slower as it requires fetching the opencv is bundle): https://lh3.googleusercontent.com/-OobUBBtnydg/X6Vn_Ls3juI/A...
Here's a demo of the addon running outside of Google Docs: https://pdftableutil.possiblenull.com/app/