Gemini can't be disabled on Google Docs
twitter.com
twitter.com
To back this up, they reference the story about Gemini chats "leaking" on the public internet. What actually happened there was that people were creating publicly viewable chats via gemini.google.com/share/ links, and some of those got posted on social media and were indexed.
If you don't want google to handle your documents, then don't put your documents in Google docs.
I personally think the two situations are quite different.
I agree that if you don’t want Google to sniff on your content you shouldn’t put it on their servers to begin with.
That said, stating that Gemini won’t remember, is dubious. Because given the track record of these companies I have my doubts that they don’t log everything they can put their hands on.
The thing you're asking it to do is to maintain the state of the document in a service. This way it's able to display the state of the document to one or more clients.
Arguably running through AI is even a more privacy-preserving feature, because Google Docs most certainly stores the data to a persistent storage (it is its main feature and cannot work without it), whereas AIs just needs to run the data in RAM.
I think using some data for something that's fundamentally different than its original use case is problematic in my view.
Another example: Google Docs indexes the contents of your document. That is, it stores all the words in a big database that you don't see and don't have access to, so that you can search for "tax" in the Google Docs search bar and bring up all documents that contain the word "tax". There is no option in Google Docs to avoid indexing the contents of a document for the purpose of searching for it.
When you decide to put your data into Google Docs, you are OK with Google processing your data in several ways (that should hopefully be documented). The fact that you seem so upset that a specific algorithm is processing your data just because it has the "AI" buzzword attached to it, seems like an overreaction prompted by the general panic we're living in.
I agree Google should be clear (and it is clear) whether Gemini is being trained on your data or not, because that is something that can have side effects that you have the right to be informed about. But Gemini just processing your data to provide feature N+1 among the other 2 billions available, it's really not something noteworthy.
Do you think this information google is gathering can then be used in the future to paginate some other document? Do you think paginating my doc will help their algorithm to better paginate documents in the future? I see what you're trying to say but putting everything in the "algorithm" bucket doesn't help moving the whole conversation around AI forward.
> The fact that you seem so upset
Your upset detector is clearly wrong. I don't use google docs. I don't care about google docs. I'm just adding my 2c to a conversation around this type of practices google and co are using.
Isn't this why we're here on HN? To exchange ideas?
“Google collects your Gemini Apps conversations, related product usage information, info about your location, and your feedback. Google uses this data, consistent with our Privacy Policy, to provide, improve, and develop Google products and services and machine-learning technologies, including Google’s enterprise products such as Google Cloud.”
“To help with quality and improve our products (such as generative machine-learning models that power Gemini Apps), human reviewers read, annotate, and process your Gemini Apps conversations. We take steps to protect your privacy as part of this process. This includes disconnecting your conversations with Gemini Apps from your Google Account before reviewers see or annotate them. Please don’t enter confidential information in your conversations or any data you wouldn’t want a reviewer to see or Google to use to improve our products, services, and machine-learning technologies.” [italics was bold in the original]
Seems pretty clear to me.
> To stop future conversations from being reviewed or used to improve Google machine-learning technologies, turn off Gemini Apps activity. You can review your prompts or delete your conversations from your Gemini Apps activity at myactivity.google.com/product/gemini.
And it is not like Gemini is the first NLP model to be reading your google docs. Google have had one of the most advanced spell/grammar checkers reading through your docs for years now.
> OK, more testing and I think I've figured it out (and it's still bad!). It seems that if you've ever clicked the Gemini button for a type of document then it remains open whenever you open another of that type--and therefore automatically ingests and summarizes it. So, e.g...
Is the concern that Google would consider your use of Gemini on this document as consent to use it for future training?
"We do not use your Workspace data to train or improve the underlying generative AI and large language models that power Gemini, Search, and other systems outside of Workspace without permission."
https://support.google.com/docs/answer/14615114?hl=en
Edit: I believe that this "without permission" caveat refers to experimental things like Workspace Labs where you have to give them that permission if you decide to join.
https://support.google.com/docs/answer/13447104?hl=en#labs-o...
I'm not sure whether there are any other beta programs that require these permissions.
1. Some people (e.g. some artists) don't like generative AI as a matter of principle, seeing it as soulless, corporate, and entirely trained on stolen data.
2. Many people resent Clippy-style popup features. They appear at the most inconvenient times, and everyone knows they're mostly for the benefit of the product manager with a user count KPI. And the harder they get pushed, the more people resent them.
3. The distinction between things-known-from-training-data, things-known-from-context and things-known-from-RAG and so on is pretty opaque to most users - and not clearly guaranteed by the documentation. If it's an assistant that can schedule reminders, can find things in your e-mails and google drive, and it promises "Personal Results" where "your communication requests will be used to improve your experience with Gemini" the distinctions are pretty ambiguous.
4. The LLM industry norm is to play fast and loose with training data.
Emotionally it makes my blood boil though. I am absolutely with the author.
I had to leave Sentry for crash reports because they did report things to me I have never asked them to watch for in the first place. Same vibe.
I don’t use Google to create personal documents precisely because I can’t trust or verify that their apps aren’t doing more than provide storage, versioning, conflict resolution, and an editing UI without my permission.
That said, I think it’s way past time to give up on hoping to trust a cloud / “big tech” provider.
There are now only three models for applications I’m happy to use, and I try to do everything in one of those three where possible:
1) Verifiably end-to-end encrypted. This is the only option way I can get comfortable with cloud software.
2) Self hosted applications.
3) Local first applications with “dumb filestore” data syncing via a provider of my choice (likely self hosted).
In all cases I have a strong preference for truly open source (avoiding VC funded open core weirdness etc.) and try to donate more than the equivalent subscription or purchase price for proprietary alternatives to the projects I find that meet my needs.
NO! Don't put your private documents on someone else's computer, that's easy.
But if you do, don't put it on the hard drive of a company that makes money from your personal data for advertising, it sounds almost childish to explain.