MemoryCache: Augmenting local AI with browser data
future.mozilla.org
future.mozilla.org
Instead of data going to models, we need models come to our data which is stored locally and stay locally.
While there are many OSS for Loading personal data, they dont do images or videos. In the future everyone may get their own Model but for now tech is there but product/OSS is missing for everyone to get their own QLORA or RAG or Summarizer.
Not just messages/docs: What we read or write, and our thoughts are part of what makes an individual unique. Our browsing history tells a lot about what we read but no one seems to make use of it other than google for ads.. Almost everyone has a habit of reading x news site, x social network, x youtube videos etc.. Ok, here are the summary for you from these 3 today.
Was just watching this yesterday https://www.youtube.com/watch?v=zHLCKpmBeKA and thought, why we still don't have a computer secretary like her after almost 30 years, who is one step ahead of us.
Local models for images are getting pretty good.
LLaVA is an LLM with multi-modal image capabilities that runs pretty well on my laptop: https://simonwillison.net/2023/Nov/29/llamafile/
Models like Salesforce BLIP can be used to generate captions for images too - I built a little CLI tool far that here: https://github.com/simonw/blip-caption
We are building this over at https://software.inc! We collect data about you (from your computer and the internet) into a local database and then teach models how to use it. The models can either be local or cloud-based, and we can route requests based on the sensitivity of the data or the capabilities needed.
We're also hiring if that sounds interesting!
We didn't do that, though. Our domain was available for like $4,000. The .inc TLD is intentionally expensive to discourage domain squatting :-)
Plus, search engines usually catch up based in click-throughs, bounces, financial kickbacks (cough), too.
Searching for Go programming language stuff was a pain a few years back, but now engines have adapted to Go or Golang.
I don't use Google, so ymmv.
I don’t think it is more useful, but it is certainly more functional (supports screen reading, text selection, maybe dark mode, etc)
might be worth having some sort of automatic fallback to a static site after a certain amount of failed loading or an error
just saw the link to your html version in another comment and it took literally five minutes to load on firefox
I don't know about everyone but a majority of searches are for stuff I've seen before, and they're often frustrated by things that have gone offline or are downranked by search engines (e.g. old documentation on HTTP only sites) or burred by SEO.
You're absolutely right about models coming to our data! If we could have Copilot-like intelligence, completely on-device, scanning all sorts of personal breadcrumbs like messages, browsing history, even webpage content, it would be a game-changer!
I was imagining something a little more ambitious. Like a model that uses our search history and behavior to derive how to best compose a search query. Bing Chat's search queries look like what my uncle would type right after I explained to him what a search engine is. Throw in some advanced operators like site: or filetype: or at least parentheses along with AND/OR. Surely, we can fine tune it to emulate the search processes of the most impressive researchers, paralegals, and teenagers on the spectrum that immediately factcheck your grandpop's Ellis Island story, with evidence he both arrived at first and was naturalized in Chicago.
That's the most important idea I've read since ChatGPT / last year.
I'll wait for this. Then build my own private AI. And share it / pair it for learning with other private AIs, like a blogroll.
As always, there will be two 'different' AIs: a.) the mainstream, centralized, ad/revenue-driven, capitalist, political, controlling / exploiting etc. b.) personal, trustworthy, polished on peer networks, fun, profitable for one / a small community.
If by chance, commercial models will be better than open source models, due to better access to computing power / data, please let me know. We can go back to SETI and share our idle computing power / existing knowledge
If you're passing in smaller documents then it works pretty good for real-time feedback.
Turns out this sort of stuff is cyclical.
I use the offline translator built into FF regularly and It's magic. I would've never thought something like that can run locally, without a server park worth of hardware thrown at it.
Here's hoping this experiment turns out the same way.
The feedback loop coming gained from chatgtp will I assume always be way better than my local gpt equivalent.
But often I bookmark pages where I know the information on there are important enough for me to come back to more than once.
So I have started crafting out a solution for this. It crawls your bookmarks on your local browser storage, downloads those pages and adds them to your search index on your OS.
That's been an itch for me for years.
Personally I would already be content if my browsers didn't forget their history all the time, both Firefox and Safari history is way too short-lived.
Google removed the feature intentionally in 2013: https://bugs.chromium.org/p/chromium/issues/detail?id=297648
Apparently Opera supported it too at the time, and from the comments Safari as well.
Performance reasons seem to have killed it. I'd think that after 10 years now any modern computer would be able to handle it.
Thinking of it, something like this can be used for all your local files as well, acting as a better version of the old filesystem-as-a-database idea. Or for a specific knowledge base (think LLM-powered Zotero).
Assuming that you can coerce the LLM to fill in the RDF correctly, and that we now have much more memory and faster storage, it might work.
Instead of doing lots of back-n-forth with the giants, enriching them with each prompt, you get a smaller local model that's much more respectful of your privacy.
That's an operating model I am willing to do some OSS contributions to, or even bankroll.
Gotta love the underdogs, even if admittedly, I am not a big Mozilla org fan.
Update:
Answering my own question it looks like it uses llamacpp in local mode? https://github.com/imartinez/privateGPT/blob/main/private_gp...
https://memorycache.ai/developer-blog/2023/11/30/we-have-a-w...
links to https://github.com/misslivirose/Memory-Cache
but did you mean https://github.com/Mozilla-Ocho/Memory-Cache
Went back a bit further/to the official site:
> MemoryCache is an experimental development project to turn a local desktop environment into an on-device AI agent.
Okayy...
And this from November
Introducing Memory Cache
https://memorycache.ai/developer-blog/2023/11/06/introducing...
> What is the meaning of a life well-lived?
Is the response to this based on browser data? Based on the description I was expecting queries more like:
> What was the name of that pdf I downloaded yesterday?
> What are my top 3 most visited sites?
> What type of content do I generally interact with?
For example, you'll see things like "where should I visit in Japan?" or "how should I plan a bachelor party?", because they are a huge variety of answers that are all "correct", regardless of how much you disagree with them. There is also a huge number of examples from them to draw from, especially compared to something as specific as your browsing history.
https://support.mozilla.org/en-US/kb/cookie-banner-reduction
https://news.ycombinator.com/item?id=38421121
The FF solution is just more automated.
https://www.russellbeattie.com/notes/posts/the-decades-long-...
(I'm suspecting because this goes against the wants of some of the biggest players who have the incentive of making us leave as many online footprints as possible ?)
Even here, Mozilla recommends converting to PDF for easier (?!?) human readability. Except PDF is a very bad format for digital documents, with no support for reflow and very bad support of multimedia. (PDF is perhaps good for archival of offline documents, even despite its other issues).
Edited to add: https://www.choosellm.com/ by the PrivateGPT folks seems to have what I needed.
I don't even like having to clear my history and wtv regularly. I use incognito mode most times.
Now I have monitor what my local AI collects?
"through the lens of privacy" my ass, man.
Why would I ask my browser what the meaning of a life well lived is?
It's important that we have Firefox working on such experiments otherwise as Google adds more of their privacy invading features to chrome / chromium it will likely impact negatively on peoples desire to find alternative browsers.
You gain market share by doing what they refused to do, no matter how much it's in the user's interest, because their business is stealing the user's data and yours isn't.