ScholarTurbo: Use ChatGPT to chat with PDFs (supports GPT-4)
scholarturbo.com
scholarturbo.com
If you have many papers, perhaps upload the latex source code of all of them into some vector database like pinecone and use them with LangChain for retrieval.
I have nothing to do with it, but to me it's the best, they seem to put some real effort on it since day one.
[1] explainpaper.com
Well, sure, but that's the file at rest. What about the security / privacy of the data as it's fed to ChatGPT? More and more of these "ask your doc / data a question" apps are popping up. Isn't this about equivalent to putting the document on the public internet or is there some sort of sandboxing involved?
I say reasonable, because in a company growing that fast, everything is on fire all the time and security is the last thing considered.
This may be a dumb question, but how would Azure ensure GPU clean tenancy and segmentation of data in these pipelines at those prices?
[0] https://learn.microsoft.com/en-us/compliance/regulatory/offe...
‘Microsoft in-scope cloud platforms & services Azure and Azure Government Azure DevOps Services Dynamics 365 and Dynamics 365 U.S. Government Intune Microsoft Defender for Cloud Apps Microsoft Healthcare Bot Service’
Are they piggybacking on that?
There appears to be a lot of apathy/total disregard to giving users any clear clue.
It might not matter to me or this audience - we probably know the direction our data is heading - but it matters to less techie people, to ordinary employees and to civil servants and government people at the bottom and the top, who have no notion where their data is processed and who they are trusting with it.
With all these ChatGPT wrappers, the hard bit is pretty much a monopolized commodity, so they better do the annoying bits really really well.
This is the second PDF chatbot I have tried, and they both suffer from the technical constraint of how many tokens you can fit into ChatGPT, meaning they refuse to answer global questions like "are there any mistakes in the document since the ChatGPT may be unaware that there is a document at all, but can answer, for example for a CV, "what jobs might you hire this person for".
This might be fixable by making round trips, for example asking chatgpt YES/NO for does this query apply to this 1000 token section of a larger document, etc. etc.
And of course if they get this right, and do it well, it becomes very useful. If it can answer based on 10000 stored pdfs, word docs, JIRA tickets, confluence pages and slack threads, then it would be super useful for organizations. It would hardly be just a wrapper then though.
1. https://archive.org/details/pepsi_gravitational_field/mode/2...
https://news.ycombinator.com/item?id=32064324 (2022)
I‘ve tried a lot of chat-with-pdf apps, including ChatDOC, ChatPDF, Humata and others. For serious paper reading, I must say that ChatDOC(https://chatdoc.com/) is the best option.
The accuracy of ChatDOC’s understanding of tables, data, and texts is significantly higher than other products. Also you can upload a folder of files and chat with the collection, easily navigating through each one.
For example: https://chatdoc.com/chatdoc/#/share/ub8Z7aarcXJHY_SzGwl3LFVw...
I uploaded a paper about GPT and asked ChatDOC to summarize the main content, explain the experimental data, and find relevant data. It completed all these tasks very well. Amazing.
He had demo videos on how to build similar applications. From what I understand the first cohort was made up of people from all manner of professions, notably consulting, finance, academia, manufacturing etc.
I knew it was only a matter of time before we see clones offering this as a service. To put it mildly, some even went as far as copying his code verbatim and using it to launch such services without acknowledging his effort.
One major downside of releasing content and open source code is the handful of users who copy and paste for commercial use without giving credit.
All of a sudden there's a wave of "chat with pdf" youtube videos using my exact content, diagrams and images.
But yes Twitter is a wild Wild West right now of people ripping ideas off any trending repos
Having spent quite a lot of professional time working through real-world problems with GPT4, I'll say it very rarely does something I'd call hallucination. It almost never invents crazy things out of whole cloth. Sometimes it makes wrong guesses about things.
You gotta keep your own head and be skeptical. Often you can tell if it's coming unhinged. If it makes mistakes and you notice, it'll definitely own them unlike many LLMs. But of course it's not perfect.
When you add parts of the content of the PDF to the prompt, which is how these tools work, then it is very likely to produce a completion consistent with the PDF.
I say “mainly” and “very likely”, but in my experience I have never had a hallucination when the prompt is augmented with factual data.
Hallucinations occur with regularity when the prompt does not contain what is requested to be completed.
I get that people have an objection to how LLMs invent things, but so do people. You should always verify things.
The core is open source (https://github.com/jiggy-ai)
I also don't use Langchain.
There will be a lot these quick to build interesting utilities to show off. There is room to feed the demand in wordpress sites and ad-word based websites.
I think OpenAI will keep the platform cheap long enough to collect the small toll and make the innovation go wild.
There is perhaps room for one more OpenAI style provider, like what Android did to iOS to feed this demand. It is a pity if its not Google.
Embeddings still seem like a more economical approach.
This doesn't make the users idiots, it just means they value their time.