YaCy: a free distributed search engine
yacy.net
yacy.net
YaCy is yet an early contributor to this research but we've yet to make something to my knowledge that gives comparable results to centralised search engines, is fast and scales well.
Google is our wonder and our curse.
That's simply not true. Things we take for granted like true multi-query search and sub-second results speeds simply aren't seen in distributed search engines.
For a while I would switch back to google when my searches failed and I wanted to check if the duck was getting me equivalent or better answers. In the past year and a half I've completely stopped. It is simply a high-quality product that exceeds google in what it offers.
Maybe you'll respond that "it just doesn't 'feel' high quality". OK, I can't argue with that because it's your opinion... but as long as we collectively decide to believe this fantasy we will continue to be locked into the search juggernaut. You can help improve the diversity of options by using something different and get to use a better product.
The only thing that I can't do on the duck is search academic papers. Google scholar is still the best thing of its type.
(Yes, you may say it has DuckDuckBot, but ask around people who run websites if they've ever seen a hit from it.)
Maybe Google makes the best bread, but it might be true that someone might make a sandwich on inferior bread that makes for a better overall snack.
And there's less than ten bakeries world-wide.
And if the sandwich maker actually got popular enough that people would only buy from them, then the bakeries would indeed give them the finger.
I think that's just a matter of habits. Initially, it feels a bit unusual and maybe even uncomfortable to use not the search engine you've got used to. But after some time you understand that now Google is the one that is uncomfortable for you.
So the average user cannot contribute to building the crawl index?
http://www.yacy-websearch.net/wiki/index.php/En:FAQ#My_peer_...
http://yacy.net/release_notes/YaCy_Release_1_90.html (Jul 2016)
My idea was using a giant P2P indexbase. However, I read the tech page, but it doesn't explain enough. For example, for this to work, there would have to be some kind of "information routing" whereby queries within certain information domains would be routed to machines. I can't see how this is being done.
Because you can't have the whole index on each machine, therefore, how does it know where to route a given query?
Additionally, peer nodes would have to crawl to a certain criteria; either subject based, or perhaps geographically. Otherwise you'd just end up with the same stuff everywhere.
If there's some more explanation, please link.
thanks.
The following words are stop-words and had been excluded from the search: [by, to].
Results: 1) Linux-Kernel Archive: By Thread 2) Series Cinematography by 3) Key to Citations
As they are bold, "by" and "to" seem to get boosted!!
The last time I looked at this (long, long ago), it was pretty much poisoned by penis-enlargement and drug ads.
I'm hoping it's gotten better since then.