DuckDuckGo vs Google
fourweekmba.com
fourweekmba.com
If anyone has doubts because they tried it years ago, I'd say go for it again.
And since I regularly browse via a Digital Ocean VM, the lack of CAPTCHAs on DuckDuckGo is refreshing as well.
If you don't mind my asking, why? And by which method do you usually accomplish this?
Edit: Thanks for explaining the CAPTCHA issue.
How?
ssh -D 8080 digital-ocean-vm-here.tld
Now port 8080 on your local machine is a socks proxy for your vm.
Google is really tor/proxy/anonymous user unfriendly. Requiring users to solve as many as four or five CAPTCHAs (seriously fuck you google).
They're like this for a reason. They didn't implement complex detection of proxies to annoy users. Just to keep everyone out who's not supposed to use their search.
But it's actually really nice to be able to pull a result or two from ddg by crawling, or to be able to use a VPN without having to solve a zillion captcha.
The only one hurt are the users.
Have you ever tried using Google's search over a VPN / Tor / public proxy? They're even worse than Cloudflare was with their CAPTCHAs).
> Why?
Probably to avoid fuckery by the local ISP and/or government.
> And how?
I'm assuming he meant OpenVPN or something, but I might be wrong.
Do you mean using a DO VM as a VPN, or via X11 forwarding to your desktop over SSH, or actually browsing on the DO VM's desktop using VNC? I've done all three in the past as experiments in private browsing, and the latter is too laggy to be comfortable.
So I figured if I had a choice of two search engines, where I get satisfying results in both of them, and one of them doesn't track me, why go for the one that tracks me?
I'd gladly do the same switch when it comes to Gmail, but I really really like Gmail's web interface, haven't gotten over that hurdle yet.
EDIT: I also switched from Google Chrome for pretty much the same reasons.
I really like the idea of Brave, but not having extension support is a major hurdle which I hope they get over.
I'm surprised there isn't a straightforward and user-friendly Google-less Chromium browser available.
Plus, Google doesn't get to see my mail any more. Ditto for Firefox vs Chrome.
Another issue I had was that my main email address was already in Google and Microsoft's systems for a couple reasons, and it wouldn't let me set it properly when changing my email address everywhere. So I have google@ and microsoft@ aliases just to work around their account management quirks.
I have multiple domains and aliases, and Fastmail is much better at those than Google Apps ever was (I also have the free plan but you can never change your initial domain), so I'm much happier.
A few for friends and family as well. --Friends I can tell to pay for themselves if I need to, but family, not so much.
Edit: As for checking them, most don't get many emails, and every good email app can handle multiple accounts easily, so I get notifications on my phone.
Really like Fastmail.
Do they have a spam filter? If yes, how do they filter emails without accessing it?
I also agree that Fastmail is great.
I could never get over the way using GMail felt like fighting with a sluggish toy version of email compared to desktop outlook (and I'm not really a fan of outlook, either). I just assumed that GMail's interface was the best a webapp could offer: it was the best I'd seen so far and had a no-longer-deserved "who could beat gmail at webmail?" bug in my mind.
I'm kinda surprised fastmail hasn't made larger inroads among the kind of techies that will occasionally bemoan giving their lives over to google.
Hard to avoid them in that case.
GOD YES! This gets incredibly annoying when searching for version-specific information, considering the differences there can be between version foo and version bar of $DISTRO. Nothing like searching for something specific to CentOS 7 and getting zillions of CentOS 6 and 5 hits.
Google will also helpfully ignore the "Verbatim" option, and God help you if you're trying to search for multiple specific phrases.
Google is looking more and more like Alta Vista in its waning days.
The problem is, whenever you can't find something in DDG, you assume it's because it's DDG, and search again with !g. Meaning I end up requerying 40% or so of my searches using !g (prefixing your query with !g redirects your query from DDG to encrypted google)
Here's some example queries I ended up !g yesterday that had much better results in google than DDG:
reduce kendo javascript file size - For me, the useful result was number 2 in google. In DDG it was number 11 ("Only What You Need | Kendo UI Getting Started").
ptr overwatch - PTR is the test patch of the game overwatch. It's regularly changed. The correct result "Overwatch PTR Now Available - August 29, 2017" is number 1 in google, in DDG the blog post August 12th is not available (well, correct from my perspective)
I tend to find anything speculative DDG is fairly terrible at. For example yesterday I was googling about trying to identify the source of some weird animation CPU cycles I was seeing in Chrome Performance Profiler, I didn't really get to the bottom of it, but I ended up ditching DDG and doing all the queries in google because I would otherwise have ended up searching everything twice.
I'm also getting quite frustrated with the "instant answer" functionality, it's generally terrible. One of the most annoying ones is the SO instant answer that just utterly sucks, you can't see the code, it usually cuts anything useful in half, and takes up a HUGE amount of space meaning on a laptop you've got to scroll to start seeing the results. I just want to be able to turn it off, but because they don't track you they don't offer that functionality.
Also the maps one is really bad. I almost always want directions, but clicking the map takes me to a really nonfunctional, bare-bones map that doesn't have directions and then I have to click another button to actually get the directions. There's a drop down where you can choose your map type, but it definitely doesn't work as expected, I just want it to always embed a google maps instead of whatever they're doing.
Basically UX ain't DDG's strong point.
Ironically, if you DDG: "turn off certain instant answers duckduckgo" it comes up with terrible search results and no answer. If you "!g turn off certain instant answers duckduckgo", the top result is at least relevant.
I also find the entire bang thing to be a gimmick. Who wants to use !r when !g with "reddit" at the end will always get you better search results, reddit's own search is abysmal (as is !so).
EDIT: I'm being overly negative, after all, I haven't actually switched back, like I did last time I tried to use DDG. So it's definitely worth a go, but I don't think it's ready for your Mum to use. Also, I have Bing on my phone's Chrome to avoid AMP, and it's actually quite good.
And I quite like using bangs, e.g. !w for Wikipedia or !i for images, rather than having to go to the initial search results on the search engine and then clicking again.
Finally, I like that on DDG, you can just arrow up/down through the search results, and then open one with enter (or cmd-enter to open it in a background tab), without reverting to the mouse.
You know you're onto something when the #1 documentation link is a key binding cheatsheet!
(And those on older versions can use VimFx until 57 is released.)
I don't ever find what I want a few times most days and need multiple sources/queries quite frequently.
That means more privacy because uses startpage.com
You can also use google on Firefox for Android. No AMP there.
Main reason: concerns about tracking and privacy.
Secondarily: politics.
DDG so far has been acceptable. As long as they keep their political opinions to themselves, the honeymoon will continue. My love affair with Google, on the other hand, is over. :(
No advertising, strong privacy policy, based in Australia (a country with a decent privacy track record that we know of).
It's $3/month which is a decent price for liberating a lot of data from advertisers.
The spam filter is playing catch-up to Google so that's a bit of a shock at first, but you can train it up well or use second layer measures like Sanebox.
Disclaimer: DDG staff - but personal opinion.
When DDG shows an ad, is there no tracking involved in that? For example, if I were to click on an ad (after turning my ad blocker off), does the destination site not get any information about me or specifically what I searched for?
For people afraid of getting off google, you can always search something like '!g my-search', it works the same for youtube(!yt), google image(!gi), or even hackernews(!hn)
And of course the best bang is the 'I am feeling lucky' one (!), i.e.: 'hackernews !'
You are linking to a ~5 year old thread that is discussing the state of things well before google switched everything over to ssl.
I can only see this as a good thing. If I want local search results, I'll add local qualifiers like "USA", "Texas", "Houston"
!ud → Urban Dictionary (what do you mean, she's 'office cute'?)
!wen → Wikipedia English
!w.. → Wikipedia (two letter language code, e.g., 'nl')
!wikt Wiktionary
So many useful bangs.Usually if I need a certain engine I just guess the bang and it's usually supported and correct. !gm for google maps, !tineye for tineye, !wayback for the wayback machine...
Setting DuckDuckGo as the search engine for your browser's address bar means all these bangs work right there whenever you open a new tab.
!hn does something or another!ddg google
Or if you are feeling lucky:
!!ddg google
:)
edit: typo
If you're on Google as your default search engine, it's not so convenient to see other engine's results (plus you're being tracked all the time).
Sounds like they have changed a lot since then, so I'll probably give it another shot.
Now that I've learned that !a works for amazon.com and duckduckgo.com gets affiliate revenue that way, then that's how I'll search amazon from now on. Need to boost up the underdog that cares about privacy, because google surely doesn't.
'<Song name> !' or '<song name> youtube !' takes you right to it.
I'd love a feature that opens the first result on a search engine like wikipedia does by default. Biggest use case for me would be imdb but there would be others.
DuckDuckGo has that by using a backslash, a space and a search term (\ bed intruder youtube) or an exclamation a space and the search term (although I believe the former is the official way now). I use it all the time, it's like a superpower.
Now if I type "h my-search" into the omnibar, it goes to google.com/search?q=site%3Anews.ycombinator.com+my-search which gives me only Google search results from HN.
This is using google's algorithm (not the search bar built into whichever website like DDG) which is still the best, especially if you constrain it to one domain name. I also don't have to type the "!"
The prefix messes up google's text prediction. I guess no one from the chromium team is using this feature, since it would be trivial to fix.
w https://www.google.com/search?&q=site%3Awikipedia.org+%s&btn... first result from wikipedia
h https://www.google.ca/search?q=site%3Anews.ycombinator.com+%...
r https://www.google.ca/search?q=site%3Areddit.com+%s reddit
y https://www.youtube.com/results?search_query=%s youtube
m https://www.google.ca/maps/search/%s google maps
b https://builtwith.com/?q=%s add "b " to a url to find out what tech is being used
wo https://www.wolframalpha.com/input/?i=%s wolfram alpha
i https://www.google.com/search?tbm=isch&q=%s google image search
g https://github.com/%s "g upspin/upspin" to jump to https://github.com/upspin/upspin.git
br https://github.com/Homebrew/homebrew-core/blob/master/Formul... read homebrew formula
(Title is google, but works for ddg amongst other things also)
We do have some of those tools, at DuckDuckGo they are called "Instant Answers."
For example the query: "2 lb to kg" will bring you to this results page https://duckduckgo.com/?q=2+lb+to+kg&ia=answer
You can see all our Instant Answers here: https://duck.co/ia
If something isn't triggering for you - let us know with the feedback tool on the site!
Disclaimer: DDG staff.
Get that 'in last year' filter already, DDG. I think I've complained about this through feedback before.
$ units
Currency exchange rates from www.timegenie.com on 2017-08-24
2980 units, 109 prefixes, 96 nonlinear units
You have: 5 lbs
You want: kg
* 2.2679619
/ 0.44092452
You have: 100 furlongs
You want: meters
* 20116.8
/ 4.9709695e-05
You have: 100 miles/hour
You want: meters/second
* 44.704
/ 0.022369363
You have: 2 kiloisraelnewshekels
You want: picodollars
* 5.5271109e+14
/ 1.8092635e-15 > 5 pounds to kg
5 * pound = approx. 2.2679619 kg
> x^2 + 8*x = 4
((x^2) + (8 * x)) = 4 = approx. x = 0.47213595 or x = -8.472136
And so on.Also the autocorrect really throws my search terms off while Google does is right most of the time.
!scholar <terms>I especially like how programming questions usually provide a top result from Stack Overflow, a lot of times I don't even need to click the link to see the answer I'm looking for.
The privacy aspect was the driver for me, but it wasn't enough to make the switch all this time. The privacy aspect coupled with an easy way to get google results when I need them is. While I've known they had the bang queries for a long time, it was actually a youtube search result that finally made me shift. I forget what it was, but I was searching for something for my kids and it looked like my activity had polluted the results. That has been a problem for quite some time and this particular instance was enough for me to make the change. It was a harmless, but I really don't like the idea that my activity will potentially bleed into that of my family, and frankly I get irritated when their activity bleeds into mine. DuckDuckGo doesn't replace youtube search, but it was more of the general principal. I tried working with different accounts over the years, but it's not easy to switch on all platforms and if you've ever tried entering a 60 character password on a playstation, you know why that's a non-starter.
I also tried it a few years ago but was disappointed by the speed - at a time when the google search responses came back "instantly", the latency of DDG was noticeable and not good-enough.
Fast forward to 2017 and that appears to have been solved and the results are as fast as google are from what my brain/eyes can tell anyway - I would not be surprised to find out that google was faster if you timed it. In the past 6+ months I've been using DDG, I've only ever found the need to switch back to google once (and that was image search)
https://www.google.com/search?q=dcss+branch+order
https://duckduckgo.com/?q=dcss+branch+order
This isn't a case where I _know_ I only want 2017 results, and so I do the syntax to filter it down automatically. I want all results, but I want to be aware of the timeline of whatever I'm going to click.
But to take the thought further: I can understand when a date isn't important. Say some documentation for a specific programming related thing. You'll probably learn to use !clojuredocs or something.
What about outside that? Those searches I can't quite describe without thinking, but my example above sort of works nicely because that game in particular has changed a bunch (and will continue to) over time and you do care about the date of a forum post or whatever.
For all I know, the answer is "that's when you use !g".
Information is here: https://duck.co/help/results/sources
Others here are saying that they need this for researching programming-related things. I just add the version number of the programming language (or whatever technology) I'm working with to the search query, when I find that the search results are outdated.
I still use !g on ddg but its about 1-2 times a month. As another poster said, I tried it a few years ago and had to quit but I tried it again a year ago and haven't switched back since
This was true when Google was created.
No one had the processing or memory available on their desktop to search an entire index of the "useful" web.
Not anymore.
How large is a "useful" index of the web today? And can it fit on your laptop? The answer is yes.
Can the entire thing be queried fast? The answer is yes.
As an example take the entire stackexchange and wikipedia dumps in their entirety(including images). Compressed it comes to 50-60 GB range. Think about that number. That's an rough approximation of all known human knowledge. It's not growing too fast. It has stabilized. To query the content you need an index.
So how large is an index to a 100 GB file? Generally around 1 GB. Let's say you use covering indexes with lot of meta data and up that to 5GB to support sophisticated queries.
With today's average hardware you can search the entire thing in milliseconds.
So why aren't we building better local search?
Because everyone is conditioned to believe, thanks to Google's success, we need to do it online. Which means baking in the problem of handling millions of queries a second into the Search problem. Guess what? This is not a problem that local search has.
Every time a chimp or a duck needs to build a protein in it's cell it doesn't query the DNA index stored in the cloud. Instead every cell has the index. Every cell has the processing power to query that index in the nanosecond time scale.
The cloud based search story is temporary.
If you want to index every reference to Taylor Swifts ass that every teenager in Norway, Ecuador and Cambodia are making, then yes you need a Google size index. But for useful human knowledge we are getting to the point where we don't need Google scale.
If you don't believe me look at what is possible TODAY in Dash/Zeal docset search for offline developer documentation search or Kiwix or with Mathematica.
Is there a way to easily sync search repositories, like Stackoverflow, Wikipedia etc., to your local computer with an automatically built search index?
I don't have a CS background so I'm not sure, is this a reasonable solution?
But having a way of making sure that a certain topic or genre of sites is well indexed locally would work. I.e. "my" top 100k. Oh, how awesome would that be.
I use https://chrome.google.com/webstore/detail/personal-blocklist... but it should come with pre-compiled lists, like uBlock Origin.
This feels like an application of the 80/20 rule. I might not always find what I need in just that offline index, but the times I do would seriously disrupt google.
Oh, I wish there were a way to rate the results in a opposite way. Maybe it's just me, but sometimes I want even mark some site as 'a hidden gem'. Usually it's something rather niche though, like a bunch of great little articles about anti-aliasing filters design and sampling, full of engineering wisdom from the decades of experience. So I think it would be perfect to have a way to tag it by some topic.
I've heard that a key early advantage of YouTube was that uploaded videos appeared immediately, not after a long processing delay. That helped it become popular, and once it was popular, it stayed popular and crushed the competition due to the network effect.
Relevant discussion: "Wallabag: a self-hostable application for saving web pages" | https://news.ycombinator.com/item?id=14686882 (July 2017)
>That's an rough approximation of all known human knowledge.
Are you taking the piss?
the problem is this would be completely useless for current events. but querying sites you use regularly, this could be an option. I often end a google search with modifiers like "wikipedia" or "reddit" or "stackoverflow". If I could store indexes to those sites offline that might be useful.
The creator of this offline search could potentially make money by indexing shopping sites like Amazon, eBay, etc with affiliate links.
You vastly underestimate index sizes. I'm not saying it wouldn't be realistic to download and index locally but they'll be much larger than 1-5%.
I get Baby Mama as #1 as well, but at least, from the movie synopsis, it doesn't appeat to be an irrelevant result.
Edit: Oops turns out its 11 years old. Better make that the last two decades then. I'm gettin old
Instead, I am wondering why there isnt a federated open source search engine. It would be a cloud of nodes, each node spidering and indexing a small hash bucket of urls. With a million such nodes we could have a live updated, distributed search engine to replace Google. We could run millions of queries without paying someone - we'd pay back by serving as a node, just like in BitTorrent. With all the interest in privacy, I wanted to see more discussion of replacing Google with an open, non-censured and private protocol for search.
Not sure if federated by they call themselves decentralised: https://yacy.net/en/index.html
Laptops generally come in 128GB or 256GB capacities. So we're talking roughly 20-50% of an average laptop's storage just to hold the search index for wikipedia alone.
Meaning it is not feasible to store a useful index of the web today on a laptop unless you want to dedicate that laptop to pretty much exclusively searching wikipedia's english content.
You're not searching that database in milliseconds on average laptop hardware, either, but you probably could get it to be "fast enough" for practical usages if you could somehow solve the harder parts like storage size, freshness, and indexing more than just wikipedia's english content.
Your indexing performance would also sharply decline when subjected to the random access seek times of those 2.5" spinning metal.
Offline search is the equivalent of the Internet Yellow Pages from back in the day. New, relevant information is being added continuously. The search index 10 minutes ago is different than it is right now.
Your reply seems pointless, yes I could download an entire copy of Wikipedia, compress it and index it - why would I want to? Are you going to do this on every machine you own? What about the updates Wikipedia receives every minute?
You should start a company though to focus on this, maybe you could call it Encarta or something?
Some say the DDG bangs are a solution. What do I win by doing that? It only made me resort to !g all the time, because the results were so bad.
Now I use https://www.startpage.com/ with region set to Swedish. It's practically a proxy for Google search, so it gives me the right results but sans the filter bubble experience (yes, I want the regional bubble). If you're a non US user, I can recommend it.
Indeed, due to the bangs, I know use DDG as my proxy to a lot of other search engines, like Wikipedia (!w/!wda/!w..) and Wiktionary (!wikt), and even obscure ones like Memory Alpha (!memoryalpha). Very handy indeed.
I rarely use region specific DDG, but when I do, I find it ok. With DDG I can manually enable/disable and switch the region, while Google guesses based on my IP - very annoying, given that I frequently use VPNs. (Of course, if you're logged in to Google, you can specify region/language/etc., but if you don't want to be tracked and delete their cookies or use the browser in private/porn/incognito mode, then it'll just assume that you are looking for Dutch things in the Netherlands, because that's where the VPN exit is...)
As for bangs, I love skipping clicks. Most of the time I know what I want and can use bangs accordingly. I'm guessing my most common tags are !a, !yt, !gh, !wiki, !wolf and then some game(s) specific bangs.
BTW, for wikipedia, !w is enough. Wolfram Alpha is !wa.
StartPage doesn't seem to be helpful at all. It doesn't let me select language and region separately. I need to select English UK and then it sets my region as UK. Also, it's very slow (at least for me)
- Google huge, DDG tiny.
- Gabriel Weinberg serial startups all failed until he sold one for $10MM, which allowed him to do and focus on DDG.
- Privacy a big deal these days and DDG markets towards that.
- DDG makes money from keyword advertising and affiliate revenue.
- You too can profit from something!
[1] https://en.oxforddictionaries.com/definition/solopreneur
This is Google marketing and brand perception at work because Google results of late, 3 years, have been unimpressive and you have to sift through pages of useless links and content to find any relevant information beyond the usual suspects one already knows, so their intensive spyware operations doesn't seem to help search quality.
It's surprising there are not more experimental search projects. One would have expected a steady stream of regular attempts but not a single credible effort exists. There was once an alternative search project called Cuil that just seemed to fizzle off.
(Actually, I suspect that a major ingredient of Google and Facebooks "algorithms" are really human reviewers... they want to keep this fact secret since human reviewers are responsible for judgement, whereas algorithms still count as impartial. This is not true of course, but it is the perception.)
Proxied (therefore non-personalized) Google results: https://www.startpage.com/
Heavily privacy-focused, even proxying content: https://www.ixquick.eu/
Meta search engine with very feature-rich search, self-hostable: https://searx.me/
Search engine with very feature-rich search result pages: https://www.qwant.com/
Not yet released, peer-to-peer search: http://pearsearch.org/
DuckDuckGo isn't even a full search engine. They don't crawl the whole Web. The heavy lifting is done by Bing and Yandex. That allows DuckDuckGo to have coverage without much infrastructure. That's what makes the business possible without too much expenditure.
Yes i have this often, and then when i click for example from page 5 to page 6 it suddenly says NO RESULTS and im always left flabbergasted with the thought "But google.. you just told me i got 11 million results to search through myself..."
Some sites popup dialogs and even if you delete all the elements they have set some of the elements in the original page to not allow scroll. With reading mode you just bypass all that crap.
Though it's possible to do a hell of a lot of site fixing in Stylish via CSS edits.
In Chrome, you can right click and select "Inspect" and then remove the relevant content from the display. :)
Other browsers have "reading mode" which tends to work well for this type of thing.
Several times I've come across an oddly formatted page and worried it'd lose styling required for the correct interpretation but not yet!
javascript:%20window.location%20=%20"http://outline.com/"%20+%20window.location
DuckDuckGo, on the other hand, is very very much worth your time, if you aren't familiar with it: https://duckduckgo.com
The one silly thing I miss about not having DDG at work: in DDG, I can type "new guid" and it gives me a new random guid. If there's a way to do that in Google, I haven't figured it out. (And yes I know there are a million other ways to get random guids. It's just convenient for me to get them this way.)
Then I decided to just look up "random 1 5", so that maybe a webpage which could deal well with intervals would show up. And there it was, DuckDuckGo understood immediately what I wanted and just generated one as an Instant Answer. Changing the interval was a matter of changing those two numbers in the query and getting the next random number worked with F5. Could not have wished for a better tool for the job.
Do you know what firewall system your work uses? We were blocked in a few of them but found that reaching out to those companies got it fixed.
Would be happy to help investigate!
Disclaimer: DDG staff.
I love DDG and what they stand for, and I'll gladly trade the creepy personalized results for the more organic results I get there.
All you need is a keyword. For instance, let’s say I’m looking for a new computer, I insert the keyword in the search box “new PC” and all you have to show me are ads related to that. I don’t necessarily have to see all the things I’ve been looking for in the past.
Seems eminently sensible to me. And much more likely to produce ads that are relevant to what I'm looking for right now.
1. DDG competes directly with Google in Google's core market
2. Not being evil is DDG's differentiator for real, they're not just paying lip service
3. DDG is small and profitable... in other words, they're a successful startup.
It can be done. It is being done.
It's my default search engine and without fail when a non-tech colleague is over my shoulder for an internet search they bust out laughing at the name and are insistent that it's a prank website. I persist and calmly explain. They relent and give that look you give a crazy person you don't want to argue with.
They also had the luxury of entering the market before there was a dominant player. Not to mention the brilliance of the pagerank algorithm.
In the past I would switch to it when I'm feeling some google-morning-after-shame (e.g. after seeing some targeted ads), I would stick with it for a day or two, but eventually go back to the 'what I wanted is on the first page' magic of google.
I've used DDG more consistently since changing the firefox search default, but there are still some things that I end up googling - sometimes DDG shows too many irrelevant results on the first page.
I think paying more attention to the bangs and moving away from 'keyword' searches will probably help - after all a tool is more useful if you learn how to use it properly - but for some topics (e.g. Haskell examples) if the first answer it finds isn't what I was looking for, the next couple of pages of results are usually useless too.
Can we please stop with this trend of random websites asking to spam me with notifications? Is there some framework everyone's using that's trying to grab notification permissions at every possible opportunity? At the very least, sites should at least tell me what they're going to have pop up as a notification and give me the option to subscribe to those notifications through some link in the actual site (like what Gmail does when desktop notifications aren't enabled yet).
I understand notifications are useful for "webapps", but for a blog post they seem entirely useless, so every time I see the prompt to allow notifications I reject it immediately on the basis of it almost certainly being spam.
Honestly, my second most-used "search" style interface these days is Wolfram Alpha. Google comes in _third._
Say what you will about their respective search result quality, I found that a lot of the delta between DDG and G! could be made up by just spending time with DDG. I believe I've subconsciously adjusted both my keyword structure (I'm conscious of using "explicitly" "quoted" "keywords" to ensure they're included in my results) and my expectations.
It might be very good. But can such a system compete with one that does use search history as context?
However, I do get your point. Without that extra context, the search results can only ever only be so good.
Imho using the search history as context can also lead to undesirable results, so can funneling people to overly popular results which still can be wrong, had both happen more than enough with Google.
It's a trade-off between spending time arguing with Google if I really meant to search for that word, regardless of " usage, and having to look at more results in DDG until I find the one that I was looking for.
DDG gives more predictable results. It reminds me of why I started using Google over other search engines decades ago.
For example, let’s say I want to look up “ear infection”. Google will spit out a bunch of info right on their results page, often saving me the trouble of even going to another site. DDG however will just give me 10 webmd links.
If they want to compete, the privacy angle isn’t enough. They need comparable functionality as well.
I know we're supposed to root for the entrepreneurs here on HN, but DDG honestly seems like more hype / marketing than a decent product. You never hear anything about startpage.com (maybe because it has a boring name?) but it just works.
If they are using a video of Snowden, they may have some philanthropic money - or it may be something nefarious. I'd love to give them the benefit of doubt though.
I wouldn't blame you, mine is too (since the majority of readers use Google), but I rank exceptionally higher in alternatives than I do on Google.
You should too.
Disclaimer: DDG staff.
Thanks for your hard work.
If the second, as I suspect, it's not really long-term viable anyway, is it.
It's an alternative to datenkrake Google that does not track you and adds some neat features (bangs, keyboard navigation, etc.)
Edit: Worth pointing out: as far as I am aware, they're using all the other search engines with an agreement with those companies, which makes it completely okay in my book to credit them as the one.
I credit my operating system to Canonical, even though the heavy lifting and behaving nicely on hardware is done by people contributing to Linux. Linux out of the box doesn't mean a thing to me. Bing out of the box doesn't mean a thing to me. DDG and Ubuntu do mean a lot to me.
That's a dangerous comparison to make because there isn't really any consensus* on which part of the stack to name the OS. For example, Linux is technically just the kernel. So most of the UX you deal with will be GNU and other user space tools. Depending on the DE you use, you might not even deal with much of Canonical's code. And Ubuntu does owe an awful lot to it's parent distribution, Debian.
In short, I get the point you're attempting to make but unfortunately your comparison is more contested than the original subject you were discussing.
* though many do advocate GNU/Linux
Canonical did a good thing of mashing it all together into a single product. DDG does the same for search in my book.
They've both outsourced the heavy lifting to some third party.
In DDG, Bing/Yahoo/Yandex does all of the heavy lifting. In Ubuntu, it's the Linux kernel. They would be nothing without those products they depend on. But they bundle that with like hundreds of different, smaller in scale products to offer a nice experience out of the box.
For smaller in scale companies that are trying to compete with big players (like Microsoft and Google), this does seem like a winning combination in my book.
But the most important feature is that DuckDuckGo is the search engine that doesn't track you. We don't store your personal information and we don't sell you out to advertisers.
Our goal is to raise the standard of trust online, and so at DuckDuckGo you are not the product, search is. We think that's an important difference.
Disclaimer: DDG staff.
So... who's paying?
Case A - if you type in "Car" we show an ad about cars. Advertisers are paying to appear at the top of that page. They do not know who you are or what you previously searched.
Case B - if you type in "Car" you are shown an ad about cars. Advertisers are paying to target based on your interest in cars, and possibly other websites you've visited, online actions you've taken, re-marketing pixels from third-party websites, or your affinity to other demographic groups.
In "Case A" the search page is the product.
In "Case B" you are the product.
Do you have specific contracts that allows you to use their indexes anonymously or are you at their mercy?
Any evidence to suggest that most people care? Most people just want the best search results, quickly.
Or is this a conscious decision to keep DDG serving a 'niche'.
> and so at DuckDuckGo you are not the product, search is.
Is DuckDuckGo a paid for product then? Users pay to search do they? If it's ad (Or VC) supported, then the user very much IS the product.
But is that information collected by you or anyone affiliated with you or given to anyone?
How about information that doesn't directly identify you, but could be used to identify you? (things like browser fingerprints, IPs, cookies, etc)
https://duckduckgo.com/privacy
We do not collect or share any information that can be used to personally identify you.
DDG's bangs are effectively a CLI interface to web searches. !g does google, !yt does YouTube, !w does Wikipedia, !wa does Wolfram Alpha, !scholar does Google Scholar...
Huge timesaver.
Having DDG with its bangs I use them all the time, and if there's some specific new site that I want to search there's (so far) always been a fairly guessable bang already set up for it.
I find it exceedingly convenient.
What prevents Yahoo from (a) offering a competitive search to DDG or (b) Yahoo from shutting of DDG from using its search engine?
I thought they where re-implementing search.. Are they "just" aggregating?
I mean, even if they were literally proxying to anonymized Google results, wouldn't that cover the privacy benefit?
Also the system in place for aggregation isn't just Nginx, it actually does things, like bangs.
Related, my biggest gripe with the WWW today is how difficult it is to find personal websites among (what I'd call) all the ad-revenue garbage.
It is an extra step to make google aware of you.
"strace" "hex" "ascii" site:stackoverflow.com
on DDG lite: https://duckduckgo.com/lite/?q="strace" "hex" "ascii" site:stackoverflow.com&kd=-1
the returned results only show "hex" and "ascii" but do not include "strace". I expect results with all three search terms (and if that fails, to return nothing).Is there anything more recent?
[1] http://highscalability.com/blog/2013/1/28/duckduckgo-archite...
Still thin results an a few sites, Ello and Reddit comments particularly IME (I use and search both frequently).
For console, there's https://duckduckgo.com/lite (no JS).
Hit counts and date-bounded search stills sends me to Brand G.
Showing ads according to a search term is totally possible, but having no attribution attached to a "click" or to an "impression" give very little advantage for the marketer who is paying for the ads. There are cost models allowing to pay for the actual action (like installing an advertised application) not just for clicks or views. Re-targeting people who had expressed their interest in a product is a useful tool for marketers as well. Having some kind of link back to the advertising campaign, which your users came from along with their LTV allow you to measure campaign productivity, which helps optimize future campaigns. And much more.
I really like the idea of not being tracked on the Internet, but it's seems like currently it's not feasible to remove a tool many marketers get used to.
This allows you to lets you share your feedback about the site or specific results with us. Your personal information is not recorded.
If a query isn't returning as expected you can share it there!
Disclaimer: DDG staff.
I started using alternatives, for privacy, but also because they are often better. Google Maps can not even find near pharmacy in EU country.
edit: right corner
If Putin himself showed up to Yandex HQ and demanded a full data dump of every search you made, he is still incapable of touching you, whereas J. Random Agent at the FBI can throw you in a 6x8 steel box.
I might describe that as gaining on Google, but still losing quite badly.
I'm still impressed by DDG of course and welcome some alternative to Google, which is a dangerous monopoly with a search that I've come to despise in recent years due to its ignoring what I've actually typed.
Every now and then when I'm not happy with the search results, I use the "g!" Shortcut to get the google results, but the occasions I'm doing this is definitely decreasing.
> "Google would be an elephant while DuckDuckGo is a mosquito (this is not to emphasize; I’m actually making things better for DuckDuckGo)."
Actually the Author is not making it better, in fact worse. From the numbers in the OP it appears Google is about 1000 times bigger (roughly) than DDG.
A better, and perhaps more apt comparison would to compare Google's Elephant to DDG's Badger (small badger).
Had a cool instant answer integration I wanted to contribute, that I hope would be useful to a lot of people (primarily students/scientists).
Seems a real shame for there to be no route to community contributions in the future.
Overall, super happy with DDG.
Quick tip: Rather than falling back to Google, try the "!sp" bang command for StartPage, which crawls Google to supplement it's results.
It appears as if Google has infected your internet.
But I understand, I try to think of the internet as a non-privacy place.
Also, the HBS case study prose style is annoying enough in itself. To have the style emulated by a non-native writer is fun. And weird.
So, for the case of web searching, you would have to know/read the content on the entire web. This is the "crawling" process that search engines perform. They read it and store/index what they need for searchability. It is entirely infeasible, obviously, for anyone (let alone every person) to do so on their individual devices.
It can perform a search without going to Google's main search page. It may or may not open the results in a browser window, though (I don't know because I don't use it). But regardless, it goes through Google's main search system, much like the proposal you made for Apple.
Whether or not it is accessed through the site directly is kind of irrelevant. It's still accessing the same remote system in the same way. Whatever interface you choose to use is just a user preference, really. You're still submitting a query to the site/service over an internet connection and you are getting the response back with the results.
Does that answer your question? I'm not sure if you were asking if this hypothetical from Apple could avoid using Google or if you were asking something else. Because if you are trying to avoid giving your data to Google, then you should also be concerned with Apple.
>I'm not sure if you were asking if this hypothetical from Apple could avoid using Google
Yes, that is what i was trying to get at!
>Because if you are trying to avoid giving your data to Google, then you should also be concerned with Apple.
Definitely!
>Whether or not it is accessed through the site directly is kind of irrelevant. It's still accessing the same remote system in the same way.
I'm just curious if there's a business opportunity for Apple here. They could basically make mobile safari switch to the hypothetical "Apple Search" for queries made in Safari's address bar (I personally don't search from google.com). It would of course only be worth anything to the end user if they saved time not being routed through google.com
In a similar vein, I've read some comments describing "Siri" as a way to get a bite of the web-search-cake.
[1] https://www.ixquick.eu/ [2] https://www.ixquick.eu/eng/privacy-policy.html#hmb
Would be a great search engine if it was the year 2001.
Why would this guy go out of his way to say this metaphor is actually literally informative and generous towards DuckDuckGo while apparently having exerted no mental effort and being incorrect by orders of magnitude on his own data?
If you're most generous to the writer, you get 12 million hits /day vs 13 billion hits /day according to his data from Wolfram, for a ratio of ~1/1100, which applied to a pygmy elephant of 5500 lbs yields a corresponding weight of ~5 lbs for the mosquito, or over 2,000,000 mg, versus the average mosquito weight of about 5 mg. Even if he meant a small 2000 kilogram pygmy elephant and a 20mg elephant mosquito, he's still 5 orders of magnitude too generous towards Google in mass comparison in his metaphor in which he is "actually making things better for DuckDuckGo".
I think most people think of animals like African Elephants when they hear "elephant", though, in which case you're looking at a 13,000 pound animal versus a ~12 pound one if you use hits as your metric, or a ~95 lb one if you go by visits as your metric.
So an actually fair metaphor is if Google's an elephant, DuckDuckGo is somewhere between a goose and a hyena. Better watch out, Google.
We also compare CPUs by die-area, not by their length or height.
Yeah, be really scared MSFT, two geeks from Stanford that wanted to sell for $800k are going to challenge you in the future.
A little (long-overdue) help from anti-trust authorities and Google might just have to watch out. After a while migration to another SE is logarithmic. https://duckduckgo.com/traffic.html Google pays for a lot of its traffic, see Traffic Acquisition Costs, directly and indirectly.
Though looking at size, an african elephant is 4m tall, while a small mosquito is 2mm. this is a ratio of 2000:1.
In which case he would actually have been generous to DuckDuckGo while not being wrong by orders of magnitude.
Conclusion: OP meant linear size, not mass. He could have stated it explicitly though, because one would naturally think that he was referring to mass.
[0] http://www.metzerfarms.com/DuckBreedComparison.cfm the most common duck (mallard) is 0.72–1.58 kg, the most common farm duck is 2-2.5 kg, but certain breeds are close to the right interval.
[1] https://en.wikipedia.org/wiki/List_of_heaviest_land_mammals
[2] https://stackoverflow.com/questions/4446112/search-for-inter...
[3] http://www.akc.org/content/dog-care/articles/breed-weight-ch...
Not really, but "exact same end result" is far from the current situation. And I'm not saying they shouldn't do it, rather that they really need to catch up.
Google (not counting other Alphabet companies) does a lots of things. I think the original analogy is mostly right.
My point can be rephrased as "given that the result is the same, do employee counts matter"? If not, then they don't matter in general.
It might tell us how realistic scaling to the size of Google is, for example. I don't think it does, but it's an interest thought: could DDG manage Google's search traffic with ~1,000 employees (scaling linearly)? Could it do the same with fewer than that?
the second one is doomed
Hey, it's just a metaphor.
Scales and ratios not needed.
edit: Wasn't meant to be negative ... as I could see myself writing something like that and my GF giggling at me for a week ...
> That was one of the most HN comments I've ever read.
What does this mean?
Well played.
I don't think the correctness or incorrectness of the scale of the metaphor really changes the way the whole article reads. Yet it's the first discussion piece I see in the comments on HN, and that doesn't surprise me a bit.
That said, it's a known quirk of the site and not a problem in my mind. I don't hate the tendency even it does cause a regular eye roll from me.
I regularly do this (to the frustration of those around me) and I think there is a strong correlation between enjoying hacker news and having this personality trait.
Okay, DDG is a hyena and not a mosquito. Great. Tell me how that changes the meaning of the article, or the strength of the facts used, or its conclusions, or the credibility of the author.
It means comments around here have a tendency to be overly pedantic and critical.
HN comments suffer from these and more.
What's this about a goose and a mosquito? What's that adding to the communication? Seems pretty irrelevant.
But the author said, specifically "I’m actually making things better", which isn't true.
You know, that's just a gratuitous swipe. So was this in the blog post: "First, DuckDuckGo didn’t start as a nerd attempt to find the ultimate algorithm."
It really looks like you want to put down nerds to make yourself feel more sophisticated or popular. This isn't high school, though, and you're pissing in the same pool you're posting in. Turning "HN" into an adjective with a pejorative tone isn't going to win you any friends on HN, especially as a new poster. It's just mean spirited.
Sometimes, taking the literal meaning might even distort the original message, like in this case changing the mosquito with a goose - if that had been the original expression, everyone would have wondered why the author chose such strange comparison.
I was going to suggest that a good author would then to choose an animal about the size of the goose that has the same characteristics as a mosquito, but why? A mosquito is considered annoying, harmful and parasitic – if it was the size of the goose, things to be very different. It would be considered a terrifying predator of the jungle. So I don't see what qualities of a mosquito would make it an appropriate choice, and this is what makes the metaphor worth correcting. It is somewhat of a slur against DDG, more so since the numbers are so dramatically inaccurate.
A good animal choose would be a… duck?
It's a metaphor. Not a fact.
English != Math :)
Video if you're unable to visualise - https://youtu.be/o0u4M6vppCI?t=2m42s
I don't notice people wearing both hats at the same time much.
Of course, the OP could be satire. When you can't tell the difference between satire and your platform, then ...
Incidentally, if we look sideways a bit the metaphor fits like: Google isn't just linearly bigger, it has nonlinear increases in advantages because of its size, clout, network effects, and so on. So the order of magnitude discrepancy is perhaps justified by attempting to account for these affects. Conversely, it could be suggesting that smaller disruptive and startups ought to spread their influence virally, parasitically, or through insect-like nimbleness and hatching-of-the-1000-eggs reproduction, instead of relying on slower, more "mammalian" reproduction to propagate their influence. Who knows? Hard to say which interpretation is correct, when words are ambiguous and one leaves it to the imagination. You can't even rely on Occam's razor to decide, since the shortest path between two thoughts differs depending on the mind. But maybe that's not a bad thing. Speech & language is nothing more than successful miscommunication. You can't police what you don't understand.
That would make this a bit of an ironic comment, no? :p
At DuckDuckGo we are averaging roughly 17 million searches a day right now (duckduckgo.com/traffic.html)
Google says they did 2 trillion searches last year[1], and while usually their search metrics include Youtube, Google Maps and Gmail queries let's assume it was all searches. That'd be roughly 5.5 billion searches a day.
So now we're talking ~1/322 or just shy of 40lbs. That's roughly the size of a Hamadryas baboon or about 22 Mallard ducks - not a bad little flock! :)
Disclaimer: DDG staff. Opinions are my own. No animals were harmed in the poor construction of this analogy.
[1]http://searchengineland.com/google-now-handles-2-999-trillio...
Powered By Reporting is listed as any search that is handled by Google or Bing. "Google’s “powered by” share is composed of searches conducted at "Google entities", as well as searches on AOL and Ask’s MyWebSearch "[3]
In Comscores qSearch product that you can subscribe to, when you hit "Google Sites" it provides a drop down detailing the breakdown of all the properties. Search is hard to define when companies get to that size, and searches on other Google properties are indeed 'searches' but they are also different than searches on DuckDuckGo.
[2]http://www.comscore.com/Insights/Rankings/comScore-Releases-... [3]http://www.comscore.com/Insights/Blog/comScore-September-201...
Today, it's just Google. I've been using DDG for a few years, but about 1/3 of the time I add an !g because I don't find the results I need on DDG.
The cost of entry to the search market is exceedingly high right now. This is a pretty good article of detailing how one person was able to come up with an idea and challenge the behemoth in very niche areas (privacy/the nsa leaks were probably the reason I started looking at/using it around 2013).
Yet I still miss the days of using multiple search engines; seeing a variety of results. I hate the de factor standard of Google. When a company controls that much of search, they get to define the narrative. They literally shape the way many people perceive the world.
I wonder if tech will get to the point where indexing will be easier and we'll see more solutions that are cheaper and that can crawl larger datasets with lower processing requirements. Maybe the next step will be distributed search with shared indexes?
In any case, Google can't remain on top forever (at least I hope not). It'd be nice to see more tech in this space, but it's an incredibly difficult problem. There is reason Google climbed to the top like it did.
Google climbed to the top because they focused on what matters: the right results, quickly, no fuss. Their competitors gave you unrelated spam for most queries and were slow and focusing on useless things.
Googles results are getting worse, there's more fuss than ever and it's getting slower all the time. So there definitely is an opportunity (which DuckDuckGo is taking).
> Today, it's just Google. I've been using DDG for a few years, but about 1/3 of the time I add an !g because I don't find the results I need on DDG.
If DDG folks are reading this, here is an idea: Have a browser extension (so opt-in). If a user then does a !g search, forward the search terms and the search results to DDG and add them to your index. (I wouldn't go as far as Microsoft, who IIRC actually copied the results and the ranking for individual terms and added them to Bing.)
Why does DDG, like Google, by default instead prefix all results with a URL pointing to DDG servers enabling DDG to track what results that the user clicks on?
Google started doing this some years ago, e.g., all the URLs in search results are prefixed with something like https://www.google.com/url?q=
Needless to say, this serves no useful purpose for users and it showed the direction the company was moving in.
Easy for any nerd to remove client-side, but it is on by default and this will catch many non-technical users who do not know about them or how to remove them.
In DDG the prefix is something like https://duckduckgo.com/l/?uddg=
Even assuming no logs are kept with client IPs, and no correlation can be made by any party in possession of DDG's logs bewteen client IP and clicked URLs, this practice is still collecting data. The data collected is not merely what searches a user submits i.e. search data, but also data about the user's browsing, i.e., which URLs she chooses to follow. And of course it collecting this data without asking for the user's permission.
If the company can argue having this data is useful to the company and therefore somehow useful to users because company will make better website/software (a common argument made by many web companies caught collecting user data by default), then it seems the respectful thing to do is ask users if they want to contribute such data. Let users make the choice.
As I recall this default prefixing of results is not something that search engines before Google used to do. Nor did the original Google do this either. It is something that Google started doing some years ago.
Note: Some browsers allow the user, e.g., to preset a default HTTP referrer, to set it to the target URL, to send an empty one, or to not send this header at all. Same options as offered by "meta referrer" except it is controlled by the user in their browser, not via a third party website. I have no idea if the popular browsers have such settings.
Finally, a problem with having bad design choices (such as prefixing URLs) and then offering users the option of changing the "settings" via a website is that this usually requires Javascript or cookies, because as everyone knows HTTP was designed to be stateless. Cookies and Javascript are two things privacy conscious users may want to avoid. By default DDG does not set cookies or require Javascript. (Good.) But if in order to change a bad default, the user has to enable cookies or Javascript, then we have sacraficed the goal of no cookies or Javascript required. (Bad.)
edit: screenshot: https://cl.ly/0W1O013N302b/Screen%20Shot%202017-09-20%20at%2...
They explain it here: https://duck.co/help/results/rduckduckgocom It's to prevent sites from seeing what search term you found them by.
Turning it off: Menu button at the top right of search results -> other settings -> Privacy Tab -> Redirect (Toggle)
Edit: Note that I have the redirect setting turned on, and don't get the redirects. Presumably because I'm running firefox nightly and they know that this browser supports rel="noopener" (which they use on the link).
But when DDG improve the search result, I will make a switch there.
1. Root the search servers. Gradually leak out what people were doing in a way that looks like fake search results to attackers running "searches" that are actually commands to trigger leaks. The leaked data would be stored in memory temporarily.
2. Subvert the engine to send malicious JavaScript to users that leaks their search results to a specific location. Might reduce risk of detection by first determining browser configuration for common ways of spotting leaks. Then, don't send anything to those users.
I'd guess the boxes and setups were originally optimized for speed and cost rather than isolation. So, what's security like at DDG? A quick glance shows they run...
"DuckDuckGo is coded in Perl and JavaScript with the help of the YUI Library, served via nginx, FastCGI and memcached, running on FreeBSD and Ubuntu via daemontools. We both run our own servers and have servers on Amazon EC2 across the world."
Probably not that private once hackers are involved with Perl and Ubuntu. On nation-state level with EC2. The servers and cache at least get security updates regularly. So, safe assumption is private against passive collection and low-to-medium-strength attackers.
Good for other threat models that majority of users are actually worried about, though. I did say that in original comment.
Security can always be improved but they already seem to be a much better choice than most of the alternatives.
For natural language queries, I use !g, because Google is better at this sort of thing. For local queries I use !gde, because Google has better map integration. For most everything else, I prefer DDG, because it doesn't "did you mean DickDickGo?", and it doesn't "I don't know what you're looking for, but here are five ads instead".
Ummm... yes it does. It's literally the very first point the author discusses once he starts comparing the search results. Here's a link to the relevant subsection: https://fourweekmba.com/duckduckgo-vs-google/#Search_DuckDuc...