55 karma · joined May 19, 2022
Tap into the collective knowledge of the online hive mind
Currently building ZEITGAIST:
> ZEITGAIST is a chatbot trained to answer questions about current global financial, cultural, and intellectual climate. It is unique, as its answers are informed not just by web sources but also by what people are saying on social media. It allows users to explore current global conversations in a truly unique way.
E-Mail: netsroht[at)zeitgaist.ai
Im currently measuring a pareto front in J/tok in order to set power limits of this card without sacrificing too much performance. Since we are talking about full power draw of ~480W which is fine during the day (with solar panels) but during night when the sun doesn't shine (even with a battery) I'd like to limit this a little bit.
Regarding "zeitgeist", about a year ago I built something similar called https://zeitgaist.ai which also incorporates other sources like Mastodon, Bluesky, some subreddits etc.
Also, I found this link [1] in the thread you mentioned. They seem to have implemented something like that.
But it's also about digital data autonomy. It's not just about avoiding surveillance over sensitive searches, but having control over our data's destiny. Even mundane data, in aggregate, can sometimes be used in ways we can't predict.
You can use search engines like Google without being logged in. When combined with tools like uBlock Origin and Cookie AutoDelete, it becomes more challenging for them to build a singular profile about a user, especially one tied to payment methods such as credit cards.
I genuinely appreciate what Kagi is doing, and I'd absolutely be willing to pay for their service, because if you're not paying for a service, you're the product. I trust companies to uphold their privacy promises, but "Trust is good, but proof is better." ;)
I hope Kagi introduces an anonymous access feature. For instance, it could incorporate zero-knowledge proofs (ZKPs). These are cryptographic techniques where one party (the prover) can confirm to another (the verifier) that a claim is accurate without disclosing any additional information. This is especially beneficial for authentication scenarios where it's essential to avoid sharing extra details.
To implement zero-knowledge authentication for quota API access:
1. Token Creation:
- Each month, users receive a token tied to their identity and quota.
- The token can be split for use on multiple devices using cryptographic methods.
2. API Access:
- Clients present a zero-knowledge proof (ZKP) to confirm they have a valid token and haven't used up their quota. The server verifies this without seeing the exact details.
3. Client Synchronization:
- Each client tracks its quota usage.
- Synchronization can be peer-to-peer or through a centralized, encrypted server to prevent double spending of the quota.
4. Quota Renewal:
- Monthly, old tokens expire, and new tokens are issued.
Challenges:
- ZKPs can be resource-intensive.
- Token security is crucial; there should be a way to handle lost or compromised tokens.
- The system should prevent quota "double-spending" across devices.
- If a centralized server is used for synchronization, it should operate with encrypted data.
This way Kagi would only know who their customers are but not what kind of searches they make.
Since this project just went live I'm still figuring out how to communicate that.
What's your opinion?
> Do you want to avoid LLMs answering your search? I have not seen that widely adopted at all.
It's starting to get more though. The brief answers that sometimes show up right beneath the search term will most likely get improved by leveraging LLMs.
Unlike Bing chat etc., I at least show the detailed sources with contents from web searches and social media comments that have been used to generate the answers.
> how to work with a memory module that remembers things about specific entities. It extracts information on entities (using LLMs) and builds up its knowledge about that entity over time (also using LLMs).
[1] https://python.langchain.com/en/latest/modules/memory/types/...
I think these tools will help us break out of local bubbles. I'm currently working on a Zeitgeist [1] that tries to gather the consensus on social media and on the web on general.
Whether humans are able to create strong AI is a philosophical question: While some argue that's not possible (can we be Gods?), others argue that this is the next logical evolutionary step.
Let's see if we can at least mimic strong AI when we let LLMs connect to external systems (internet, money, more energy, etc) and specifically allow themselves to fine-tune or train new NNs in general.
Time will tell.
I would never send unencrypted PII to such an API, regardless of their privacy policy.