YaCy: Decentralized Web Search
yacy.net
yacy.net
This is a huge liability for google. Any company with 98% of its revenues tied to one product necessarily creates a fundamental liability for itself. I'm sure google execs realize this, and that's why they are pouring so much revenue into "moonshots" and subscription services (not to mention government contracts). They are actively diversifying google's revenue because they realize it comes from a near single source, search, which may represent a fundamentally unsustainable product.
Can Google control the search market forever? Are people (specifically technophiles and "early adopters") not growing increasingly frustrated with them? The same groups of users that Google once relied upon as an initial source of traction are now abandoning the company's products in search of more open pastures.
Decentralization is an unstoppable trend with momentum across product verticals. File sharing was the first "mainstream" invocation of decentralization technology, and blockchain/Bitcoin is the most recent. (Interestingly, Bitcoin itself is a meta-enabler of decentralization as it introduces the possibility of an automated payment layer.)
Is it inevitable that that a decentralized search engine will replace the current centralized model that Google requires for its sustained business? When will it happen? How can such a movement gain momentum?
If I were an executive at a google competitor, I would be actively exploring these questions, and finding ways to push decentralized search products into the mainstream. Mozilla, Yahoo, are you listening?
And Mozilla has always been happy to ride the coat tails of Google Search. Declining AdWords revenue just means declining Mozilla revenue. Even if the two aren't partnering together going forward, Google helped Mozilla negotiate a high bid with Yahoo.
Long term, I think Mozilla wants to see Firefox OS compete with Android. There's a lot more revenue potential there and the best chance to rapidly grow their user base.
This blind spot in people's understanding of how the internet worked, works, and should work is why we got stuck accepting centralized services as normal. They are not sustainable though and will be superceded.
The cost of providing Google's organic search service is dropping as compute hardware gets cheaper. But the ad volume isn't dropping along with it. That creates a vulnerability for Google - their "free" can be undercut on price. The problem with having high profit margins is that, to a lower-cost provider, you're lunch.
"Ello" is trying to do that to Facebook. "Ello" is probably much smaller than they claim to be; almost all their Google hits were in the first week of operation. But someone else may be able to bring that off.
I wonder if Bitcoin2.0 projects like StorJ, Bitcloud, Filecoin will be instrumental in the creation of such a product.
Would love to see adoption of this though. It would be great if there was a web interface for it for folks afraid to / can't install a fat client. It wouldn't contribute to the network, but more people would use it, which might lead to people with cash supporting it.
I didn't see a way for anyone who wasn't operating their own server to use the service though. It seems like a lot of the quality control / model building power of google has to do with the volume of in bound queries (they're able to see what user satisfaction with results is based on clicks). It seems like having a public facing server that interacted with the peer to peer network might help with adoption from less technical users.
EDIT: Nevermind found this demo portal (http://search.yacy.de/) although it's not very prominent
The software is far away from being ready, esp. the kernel, the distributes search is not really implmented, the last months I had not much time to contribute, however, what is done is available here under an open source license: https://github.com/r10s/gosearch
On the left hand side, after a search it allows you to "refine" your search by categories. Stuff like author, site, filetype, language, with each one having a count next to it to display the potential amount of hits with that filter/category activated.
Additionally, looks like it has some sort of "stealth" mode. It appears to limit the search to your own peer, or what I'm guessing is already on your local index. That could be handy by itself, if properly configured, to a local or personalized search.
What are people's thoughts on combining social graph + blockchain + decentralized search? The idea is that your searches will be somewhat similar to others in your community, so the crawl index is sharded/partitioned to optimize for social graph proximity. If you want to index pages non proximate to you, you can get paid Bitcoin to do it.
This could be implemented with xmpp (lookup socialvpn/ipop project) for social layer, chord dht for search, and Altcoin with modified pow for incentive.
The rest of the time they were just about ok and I'd find what I wanted on the 2nd/3rd page.
It was a good way to throw up some left-field results though, and for that reason it's worth keeping a node going if you can. Might get it up on the new VPS.
We might get to a pretty efficient system that way. We see this with bitcoin: Mining is done with the cheapest power available and highly optimized algorithms and hardware.
If the advertising income flows back into the spidering - It could become kind of the perpetuum mobile of distributed websearch.
How about using this type of technology for something Google can't (for legal reasons)? Say for example full text search of the library genesis archive? Over Tor or somesuch?
http://yacy.net/en/Searchportal.html
but all it searches is the forums for the Free Software Foundation, Europe. That's a job that could probably be handled by one server.