81 karma · joined April 9, 2021
- Converts each page in a PDF document into high-resolution images - Detects texts, tables, links, and images from the high-resolution image using Vision LLMs and parses them in markdown format - Handles multi-page PDF documents effortlessly - And it's easy to get started with this library (just pip install vision-parse, and then a few lines of code to convert a document into markdown formatted content).
RAG and fine-tuning serves a slightly different purpose. Fine-tuning helps LLM in learning a new task/skill such as question/answering task, summarization task etc, and improving reliability at producing a desired output such as JSON format structure thereby reducing dependency on prompt engineering.
On the other hand, RAG provides you with external domain-specific knowledge, which one can leverage to get latest information.
I have shared a notebook and explained detailed steps in my blog - https://medium.com/@iamarunbrahma/fine-tuning-of-falcon-7b-l.... If you are interested to replicate the steps on medical-domain therapy chat transcripts, you are definitely welcome. If you face any issues during fine-tuning steps, you can connect with me on my blog. Would love to help out!