It would be neat to see this as a TOAST type in Postgres, where the PDF was kept in a data structure with the PDF parsed. It would be relatively straightforward to perform searches and index/reindex deep into the documents.
In practice though, a PDF for most cases has text-like semantics, so with ::pdf::text you can have all the text-indexes you want.