Ask HN: Is there a ready-to-go solution to parse documents content?
I'm looking for a ready-to-go solution to parse documents (such as pdf, docx, pptx and others). By 'parse' I mean text extraction, including OCR if needed. I know about Tika and tried it already, but are there any more reliable alternatives, maybe based on Tika? I'd like to interact with it via REST API.
Thnx