The challenge we faced (for the two days we tried to hack together a solution) wasn’t so much the OCR but the fact that we couldn’t standardize the form as we were onboarding people with the documents from other banks.
Our requirement was that the the software would need to read the number of separating lines in the table and output accordingly.
Maybe it can do this, idk. I was an intern at the time.
We were using tesseract for our attempt too.