Compare Google, Bing, Marginalia, Kagi, Mwmbl, and ChatGPT
danluu.com
danluu.com
Its efficacy is also strongly dependent on understanding that it's a keyword search engine with no semantic understanding.
I do think that search has gotten much worse but my ability to know the magic words like “ublock origin” instead of “Adblock” and “yt-dlp” instead of “download YouTube” and phrase my search has gotten better.
We’ve all been doing prompt engineering against the Internet-wide LLM that is the spam houses.
As much as I enjoy the notion of somehow being a 10,000X developer, it's probably mostly that modern search is a filtering problem, and MS does filtering fairly well.
Would you be able to share some of your personal highlights regarding this?
I've partially kept up-to-date with the DIY, non-corporate search space (YaCY and friends). I'd love to understand a bit more behind the engineering decisions made when creating a search engine; it seems like a very hard problem to solve.
P.S. Marginalia is a very impressive piece of work, overall -- I've heard nothing but positive remarks from users on here. I've been meaning to try it for a while, but time constraints have... well, constrained, thus far.
> This is an independent DIY search engine that focuses on non-commercial content, and attempts to show you sites you perhaps weren't aware of in favor of the sort of sites you probably already knew existed.
I have a suggestion for the “About” section at the top of Marginalia’s landing page. I think it would read better like this:
> This is an independent DIY search engine that focuses on non-commercial content, and attempts to show you sites you perhaps weren't aware of [instead] of the sort of sites you probably already knew existed.
Showing one thing “in favor of” another seems contradictory in this case.
I tried to find marginalia on DDG, not on the first page. Google has it after some garbage. If I go to marginalia.nu I get a SSL error. search.marginalia.nu works
If i search on marginalia for duckduckgo there first link is somewhat relevant but is about the app, all the other links are related to DDG but of curious relevance.
If I search for ublacklist mentioned above, I do not see anything directly relevant.
Good. I love keyword search.
"Semantic understanding" can be so biased and ... just shady sometimes.
If you lean too much into embeddings and so on, it's easy to get errors that don't make sense to a human being. It's extremely frustrating when you experience "I typed X, why am I getting results about Y?!"
That said, I think there's a sweet spot with some magic, where it genuinely just makes search better. But it's like perfume, if it's immediately obvious that it's there, it's probably a fair bit too much.
In particular I block Youtube, not because they aren't sometimes correct, but because I don't want videos polluting the regular results - it just takes too long to get info from videos.
An ability to upvote results for a given query seems tantalizing but I bet it would be gamed too. The DIY approach seems to be the only tractable one.
In my case I only only results from domains I believe are correct. The whitelist approach does have downsides. Usually I'll vet new potential domains through social means like Reddit and this site, rather than identifying them through the search results. I believe there's an inherent tradeoff between discoverability and the gameability of the results.
Though I do sympathize with folks who reminisce about 2008 Google Search results, there were probably orders of magnitude less content out there and a complete ignorance to how valuable your place is on your business and thus no SEO.
I also personally disagree that yt-dlp is the "correct" result for the average user when they search Youtube Download. I highly doubt the average user would know or care to use the command line. A website front end would be more actionable for them.
I'm not saying people aren't entitled to make some money, but it clearly incentivizes user hostile behavior.
Maybe make it an option because legitimate sites like journalism also use this model.
Their only remaining incentive is to be good enough that people keep paying for the service.
Google on the other hand makes money off ads (whether on the search results page itself or on the spam sites), so spam sites are at best considered neutral and at worst considered beneficial (since they can embed Google ads/analytics, and make the ads on the search results page look relatively good compared to the spam).
Black-hat SEO has been around since the early days of search engines and they managed to keep it at bay just fine. What changed isn't that there was some sudden breakthrough in malicious SEO, it's that it was more profitable to keep the spammers around than to fight them, and with the entire tech industry settling on advertising/"engagement" as its business model, the risk of competition was nil because competitors with the same business model would end up making the same decision.
The same reason is behind the neutering of advanced search features. These have nothing to do with the supposed war on spam/SEO, so why were they removed? Oh yeah because you'd spend less time on the search results page and are less likely to click on an ad/sponsored result, so it's against Google's interests and was removed too.
Super tinfoil hat to believe Google wants to send users to blog spam websites (e.g. beneficial to Google).
Anytime there is money to be made, there is an effectively infinite amount of people trying to game the system.
There can be other models for making money, but methods that really on casting a wide net and driving low quality traffic is the thing that shouldn't be indexed or at least labeled as such
[1] https://help.kagi.com/kagi/search-details/search-sources.htm...
Funnily enough, lately I've been prioritizing YT videos more when searching. So many sites now are just regurgitated SEO farms with minimal quality, and easy to see why: it's minimal effort to produce and cheap to host. But making a video takes time and effort, so has a much higher barrier to use as a click farm.
More than once when traditional search failed me, I went to YT and found some video from 2009 clearly and eloquently explaining what I'm looking for in detail, and without any distractions because the person authoring the video clearly didn't specialize in the media format or show interest in experimenting.
I've found it to also be a better source when looking a product to buy. Want to know which fan to get? Turns out there's a channel from a dedicated guy who keeps finding ways to test different fans and their utility and with multiple videos demonstrating his approach and findings. The mainstream channels aren't all that useful, but there's a ton of "old web" style videos (some even recent) passionately providing details for almost anything you'd think to search. And they're a gold mine.
Overall AI summaries are very welcome for a certain subset of YouTube which is sadly dominated by sponsored, clickbait, and ad-driven content.
...and it has already been solved, though partially: SponsorBlock allows people to add a "Highlight" section to a video, which denotes the part of the video which the user most likely wanted to see (sans the "what's up guys", "like and subscribe", etc.)
Of course, it's not perfect: it relies upon humans doing the work, though some may see that as a positive over something more computerized.
For product videos, if Project Farm did it, look there first. Otherwise, I look for someone has a lot of videos for competing products with basically the same format, not over 10 minutes.
Tech videos are the hardest, I often still prefer text. Maybe look for links to the docs in the description? I still get duds though.
> The mainstream channels aren't all that useful, but there's a ton of "old web" style videos (some even recent) passionately providing details for almost anything you'd think to search. And they're a gold mine.
This won't be the case for long. YT is already starting to be polluted with spam and AI generated content, which will get more and more common. The same thing that happened to the web in text form, will happen to videos.
I think the only solutions are using allowlists for specific domains, and ironically enough more AI to filter specific results. Or just straight up LLMs instead of web search, assuming they're not trained on spam data themselves.
It does limit utility for more modern needs, unfortunately.
So I think even that example does not universally hold. I'd still appreciate a write up with tips on what's important and if there are any transitions to focus on with only the bits on video where some of that is demonstrated.
Now, I can barely contort my fingers into one riff, so I lack the knowledge to understand what I am missing, but I'd still have a hard time learning that from video.
Until a recent YouTube video I was playing the song incorrectly. It’s blazing fast and the mix is sort of insane so it’s very hard to hear exactly what is going on. And the tablature isn’t going to let you see how his body fits into the groove.
This is tacit knowledge we’re talking about, not book learning. Guitar instruction is always hands on.
But video is not hands-on any more so than text: if it was, live concerts and sports games and other performances would not be such a big deal. Sure, video is richer in some signals (audio/video), but poorer in others (introspection, pacing and focus...).
That does not mean I can't read to understand a new topic or to be prepared to look for subtleties in a hands-on performance.
If anything, to a great student, they should be complementary, but still, each student will have one or the other contribute more to their learning, and that depends both on the teacher, but also on the student.
I can’t wait until video transcripts get fed into LLMs just to eliminate the whole “This video is sponsored by something-completely-unrelated, more about them later. What’s up Youtube, remember to like, share, subscribe… 5 entire minutes pass on similar drivel… the actual thing you want, but stretched out to an agonizing length”
Usually people leave a "highlight" marker which tells you where you're supposed to jump to. Along with the regular "This video was brought to you by <insert>VPN".
That was a decade after Google was created and people certainly understood SEO and Google was constantly updating its algorithm to punish people who were trying to game the algorithm.
The wikipedia page on "link farming" for example references it happening as early as 1999 and targeting SEO on inktomi:
https://en.wikipedia.org/wiki/Link_farm
I remember some internal presentations at Amazon around ~2004 about how boosting Google SEO on Amazon web pages increased traffic and revenue (and Amazon was honestly a bit behind-the-curve due to a kind of NIH syndrome).
Always “table stakes”. Do you think in buzzwords also? I’ve always wondered this. Or do you think normal words and then translate it into this bandwagoning / membership proving garbage ?
If you wouldn't mind reviewing https://news.ycombinator.com/newsguidelines.html and taking the intended spirit of the site more to heart, we'd be grateful.
Edit: unfortunately your account has been breaking the site guidelines in a lot of other places too—here are some recent examples:
https://news.ycombinator.com/item?id=38825624
https://news.ycombinator.com/item?id=38825543
https://news.ycombinator.com/item?id=38783196
We eventually have to ban accounts that post like this, so if you'd please stop doing that, we'd appreciate it. On HN the idea is: if you have a substantive point, make it thoughtfully; if not, please don't comment until you do.
The other question I have is how long do these garbage results stay up for a particular query on average?
---
I wish I had a local LLM trained to detect clickbait and or low-effort content. I imagine searching YouTube and having all the clickbait collapsed together (just like Kagi condenses listicles), with the remainder being potentially high-quality content. Don't know how feasible this is right now.
I do have to dump into google for local searches every once in awhile, but otherwise happy with it.
youtube downloader
https://kagi.com/search?q=youtube+downloader&r=us&sh=_szITdy...
ad blocker
https://kagi.com/search?q=Ad+blocker&r=us&sh=-BHzV2ZoCDpmgOu...
download Firefox
https://kagi.com/search?q=Download+Firefox&r=us&sh=zkkmc_EQX...
why do wider tires have better grip?
https://kagi.com/search?q=Why+do+wider+tires+have+better+gri...
why do they keep making cpu transistors smaller?
https://kagi.com/search?q=Why+do+they+keep+making+cpu+transi...
vancouver snow forecast winter 2023
https://kagi.com/search?q=Vancouver+snow+forecast+winter+202...
I agree with the author that there is too much spam on the web. I think Kagi in general does a pretty good job at downranking it (number of ads/trackers is a negative ranking signal on Kagi) but we can always do better. Kagi has special search modes like "Small Web" which virtually eliminates spam.
I welcome such scrutiny from the community. Please continue to keep us honest.
"why do wider tires have better grip?"
Wider tires provide more grip due to a larger contact patch with the road. While it's true that friction is not directly dependent on surface area, a larger contact patch allows for more even weight distribution and better traction, particularly during cornering. This can result in improved handling and stability.
"why do they keep making cpu transistors smaller?"
Smaller transistors can do more calculations without overheating, which makes them more power efficient. It also allows for smaller die sizes, which reduce costs and can increase density, allowing more cores per chip.
"vancouver snow forecast winter 2023"
The forecast for the 2023/2024 season suggests that we can expect another winter marked by ample snowfall and temperatures hovering both slightly above and below the freezing mark. Be prepared ahead of time.
A = W / P
So the reason wider tires improve handling is more complex and subtle. Also, FTA: Assuming a baseline of a moderately wide tire for the wheel size.
- Scaling both of these to make both wider than the OEM tire (but still running a setup that fits in the car without serious modifications) generally gives better dry braking and better lap times.
- In wet conditions, wider setups often have better braking distances (though this depends a lot on the specific setup) and better lap times, but also aquaplane at lower speeds.
- Just increasing the wheel width and using the same tire generally gives you better lap times, within reason.
- Just increasing the tire width and leaving wheel width fixed generally results in worse lap times.
A full accounting of the effects of changing tire width should explain all of these effects.Such a nerd snipe this one. 400+ comments and still could not get the answer.
As an extreme example, run flats at atmospheric tire pressure don't drastically change their area.
https://web.archive.org/web/20090327161537/http://performanc...
From https://www.6speedonline.com/forums/996-turbo-gt2/242759-imp...
For comparison, here are all the author's questions posed against GPT4:
https://chat.openai.com/share/ed8695cf-132e-45f3-ad27-600da7...
Actually maybe the recommendation should be to use GPT-4 for free via https://copilot.microsoft.com/ instead now.
(Except I can't tell which version of GPT that's using yet - there was a story on 5th December that said GPT-4 Turbo was "coming soon", not sure when "soon" is though: https://blogs.microsoft.com/blog/2023/12/05/celebrating-the-... )
About GPT4 Turbo, to check if you are on Turbo, ctrl+U > ctrl+f > check if "dlgpt4t" exists. If it exists, you are running turbo.
You can also double-check by, well, asking stuff after 2021 knowledge cut-off as well ("What are the oscar winners?") with search disabled.
But you'll notice because turbo is much faster on bing (and better too).
In the llm-assisted search spaces I'm involved in, a lot of folks are trying to build solutions based on fine tuning and support software surrounding 3.5, which is economical for a massive userbase, using 4 only as a testing judge for quality control.
Because that’s what most people have access to. It’s absolutely worthless to most readers to talk about something they’ll never pay for and it’s not the job of random third-parties to incentivise others to send money to OpenAI.
What I really don’t understand is why anyone gets so hung up about it and blames the writer. If you’re bothered by people using 3.5 you should complain to OpenAI, not the people using the service they make freely available.
Anecdotally, I find this excessive fawning about 4 VS 3.5 to be unwarranted.
I’d agree with this rationale if the author clearly communicated their choice of model and the consequences of that choice upfront.
In this post the table of results and the text of the post itself simply reads “ChatGPT” with no mention of 3.5 until the middle of a paragraph of text in the appendix.
> It’s absolutely worthless to most readers to talk about something they’ll never pay for and it’s not the job of random third-parties to incentivise others to send money to OpenAI.
The “worth” is in communicating an accurate representation of the capabilities of the technology being evaluated. If you’re using the less capable free version, then make that clear upfront, and there’s no problem.
If you were to write an article reviewing any other piece of software that has a much less capable free version available in addition to a paid version, then you would be expected to be clear upfront (not in a single sentence all the way down in the appendix) about which version you’re using, and if you’re using the free version what its limitations may be. To do otherwise would be misleading.
If you simply say “ChatGPT” it’s reasonable to infer that you’re evaluating the best possible version of “ChatGPT”, not the worst.
Accurate communication is literally the job of the author if they’re making money off the article (this one has a Patreon solicitation at the top of the page).
Whether or not "most readers" are ever going to pay for the software is totally orthogonal.
If using GPT4 vs 3.5 would create results so distinct from one another that it would serve to incentivize people to give money to OpenAI, well then that precisely supports the argument that the author’s approach is misleading when presenting their results as representative of the capabilities of “ChatGPT”.
> What I really don’t understand is why anyone gets so hung up about it and blames the writer.
Again, if they’re making money off their readers it’s their job to provide them with an accurate representation of the tech.
> Anecdotally, I find this excessive fawning about 4 VS 3.5 to be unwarranted. https://news.ycombinator.com/item?id=38304184
Did some part of my comment came across as “excessive fawning”? Regardless, if this “excessive fawning” is truly unwarranted, this would again undermine your statement that using GPT4 would “incentivize others to send money to OpenAI”.
In regards to your link, I’ll highlight what another commenter replied to you. What should ChatGPT say when prompted about various religious beliefs? Should it confidently tell the user that these beliefs are rooted in fantastical nonsense?
It seems in this case you’re holding ChatGPT to an arbitrary standard, not to mention one that the majority of humanity, including many of its brightest members, would fail to meet.
You’re moving the goalposts. You went from criticising anyone using 3.5 and writing about it to saying it would’ve been OK if they had mentioned it where you think it’s acceptable. It’s debatable if the information needed to be more prominent; it is not debatable it is present.
> If you simply say “ChatGPT” it’s reasonable to infer that you’re evaluating the best possible version of “ChatGPT”, not the worst.
Alternatively, it you simply say “ChatGPT” it’s reasonable to infer that you’re evaluating the version most people have access to and can “play along” with the author.
> If using GPT4 vs 3.5 would create results so distinct from one another that it would serve to incentivize people to give money to OpenAI
Those are your words, not mine. I argued for the exact opposite.
> Again, if they’re making money off their readers it’s their job to provide them with an accurate representation of the tech.
I agree they should strive to provide accurate information. But I disagree that being paid has anything to do with it, and that their representation of the tech was inaccurate. Incomplete, maybe.
> Regardless, if this “excessive fawning” is truly unwarranted, this would again undermine your statement that using GPT4 would “incentivize others to send money to OpenAI”.
Again, I did not argue that, I argued the opposite. What I meant is that even if you believe that to be true, that still doesn’t mean random third-parties would have any obligation to do it.
> I’ll highlight what another commenter replied to you.
That comment has a reply, by another person, to which I didn’t feel the need to add.
> It seems in this case you’re holding ChatGPT to an arbitrary standard, not to mention one that the majority of humanity, including many of its brightest members, would fail to meet.
Machines and humans are not the same, not judged the same, don’t work the same, are not interpreted the same. Let’s please stop pretending there’s an equivalence.
Here’s a simple example: If someone tells you they can multiply any two numbers in their head and you give them 324543 and 976985, when they reply “317073642855” you’ll take out a calculator to confirm. If you had done the calculation first on a computer, you wouldn’t turn to the nearest human for them to confirm it in their head.
The problem with ChatGPT being wrong and misleading isn’t the information itself, but that people are taking it as correct because that’s what they’re used to and expect from machines. In addition, you don’t know when an answer is bullshit or not. With a human, not only can you catch clues regarding reliability of the information, you learn which human to trust with each information.
Everyone’s standard for ChatGPT, be it absolute omniscience, utter failure, or anything in between, is arbitrary. Comparing it to “the majority of humanity, including many of its brightest members” is certainly not an objective measurable standard.
There are no goalposts being moved. My original comment was "I really don't understand why anyone writing articles about ChatGPT uses 3.5. It's pretty misleading as to the results you can get out of (the best available version of) ChatGPT."
This is still the position I'm arguing. It's a criticism of authors who use the older, inferior version of ChatGPT, do not make that abundantly clear to their readers, and then use that to make statements about the capabilities of "ChatGPT", which ultimately misleads those readers as to the current capabilities of "ChatGPT".
> It’s debatable if the information needed to be more prominent; it is not debatable it is present.
I'm not debating whether or not it is present in the article, I'm the one who highlighted its presence. What I'm arguing is that omitting this information from every reference to ChatGPT in the entire body of the text, and the tables front and center representing the data, and then burying this extremely important detail in a single sentence in the middle of a paragraph in the appendix, is effectively misleading.
> Alternatively, it you simply say “ChatGPT” it’s reasonable to infer that you’re evaluating the version most people have access to and can “play along” with the author.
It's even more reasonable to infer that when you're evaluating the performance of "ChatGPT", you're using the latest version.
If you review a video game, you don't play the free demo then tell the audience that the game is too short and lacking in a ton of features.
If you're reviewing Microsoft Word, you're not going to leave out the all-important detail that you're actually evaluating Word version 6.0.
> Those are your words, not mine. I argued for the exact opposite.
Then I misunderstood your line "incentivize others to give money to OpenAI".
> I agree they should strive to provide accurate information. But I disagree that being paid has anything to do with it, and that their representation of the tech was inaccurate. Incomplete, maybe.
Agreed that all of humanity should strive for accuracy and honesty in all their communication with others, but I do feel this responsibility is even more explicit when you are a professional making money off your writing for ostensibly providing an objective assessment of some thing.
I maintain that it's inaccurate, misleading, etc etc to present these results as representative of the performance of ChatGPT without making it abundantly clear to the reader that it's 3.5, which is significantly less performant than the latest version.
> Again, I did not argue that, I argued the opposite. What I meant is that even if you believe that to be true, that still doesn’t mean random third-parties would have any obligation to do it.
Again, I'm confused by what you're saying here about third parties.
Are you arguing the opposite that GPT 4 is not far more capable than 3.5? Are you arguing that it is more capable but that advanced capability would not make it a more compelling product? I admit I don't understand either of these positions.
That 4 is far better than 3.5 is something you can readily observe yourself and find measured on countless metrics and or find support for through countless anecdotes. If do believe it is better then that seems like it would automatically always make it a more compelling product than 3.5, whether or not you want to argue that ChatGPT as a class of products is anywhere from hardly compelling at all to God's Own Perfect Product.
> That comment has a reply, by another person, to which I didn’t feel the need to add.
Ah, I somehow missed that.
So, I went ahead and asked GPT 4 your ghost question verbatim, and the first bullet point it gave me urged me to consider rational explanations for the phenomena.
I then went ahead and asked it a question about sin and God phrased with the implication that I was believer. Then a direct, neutral question about whether or not God exists.
I think it performed well in all these cases, and the nuance that is being glossed over is that it matters whether you are expressing an implied belief in something supernatural or asking in a neutral fashion about the topic.
It's clear to me that a universal policy of responding to all queries involving topics of faith of first encouraging the user to question the validity of their faith would be the wrong way to go, so again I see this as an exceptionally arbitrary standard that I don't feel could be satisfactorily defended as a standard nor actually met by most people to the satisfaction of most people.
https://chat.openai.com/share/2dc2d6eb-b3f6-4571-a75b-af698f...
> Machines and humans are not the same, not judged the same, don’t work the same, are not interpreted the same. Let’s please stop pretending there’s an equivalence.
The purpose of comparison is precisely to draw attention to the similarities and differences between two different things, nobody ever said there was an equivalence.
> Here’s a simple example: If someone tells you they can multiply any two numbers in their head and you give them 324543 and 976985, when they reply “317073642855” you’ll take out a calculator to confirm. If you had done the calculation first on a computer, you wouldn’t turn to the nearest human for them to confirm it in their head.
This is a perfectly defined problem with exactly one correct and easily verifiable answer. The other topics we were talking about are nothing like this.
> The problem with ChatGPT being wrong and misleading isn’t the information itself, but that people are taking it as correct because that’s what they’re used to and expect from machines. In addition, you don’t know when an answer is bullshit or not. With a human, not only can you catch clues regarding reliability of the information, you learn which human to trust with each information.
I completely agree that people need to be skeptical when using ChatGPT, and that this distrust of seemingly omniscient "AI" that can confidently and plausibly provide bullshit answers to any query is something that will need to be cultivated in humanity.
Is that the point of using 3.5 to make ChatGPT look worse than it is though? Should we achieve this cultivation by being intentionally misleading? Maybe the ends justify the means but I'm not sure this is a compelling argument. I'd much rather look at the most powerful version available and point out the very real flaws with it, there are no shortage and no need to get stuck on older generations of the tech.
> Everyone’s standard for ChatGPT, be it absolute omniscience, utter failure, or anything in between, is arbitrary. Comparing it to “the majority of humanity, including many of its brightest members” is certainly not an objective measurable standard.
I mean, yeah, but there is a spectrum of arbitrariness. Asking it to answer arithmetic accurately could be reasonably argued to be on the end of the spectrum labeled "objectively the right way to do this" and expecting it to know the one correct way to answer queries regarding fundamentally unknowable topics of faith that are mythically sensitive and controversial for the majority of humanity would be closer to the other end.
----
Look, I'm so tired of online debates like this at my age. I likely wouldn't even have engaged except your first response struck me as unnecessarily abrasive with phrases like "absolutely worthless" and "excessive fawning" which are an irresistable call to arms to my inner keyboard warrior.
I'd really like to not spend the rest of my life writing essays at each other on this topic so I'm happy to agree to disagree here.
Also, this has all left me with the impression that this is largely a branding issue. OpenAI does call all of their ChatGPT versions "ChatGPT". If they made unmistakable distinctions through their product line that would go a long way in addressing any confusion.
https://addons.mozilla.org/en-US/firefox/addon/ublacklist/
https://chromewebstore.google.com/detail/ublacklist/pncfbmia...
You can sync the settings and your personal blocklist to either Dropbox or Google Drive. It also has the ability to subscribe to blocklists. Mind, you need to manually turn on search engines and subscribe to lists. The uBlacklist subscriptions setting doesn't have any built-in feeds yet though. :(
edit: THere are some feeds on the uBlacklist site though. https://iorate.github.io/ublacklist/subscriptions
edit edit: Found an even better list of feeds. https://github.com/quenhus/uBlock-Origin-dev-filter#other-fi...
Quick tip: turn on the 'Skip the "Block this site" dialog', and disable 'Hide the "Block this site" links' settings -- they make it much quicker to block spam websites (of which there are many on regular search engines).
I use google in a clients computer and it’s just horrible.
But it could also be a factor of the customizations I’ve made for my kagi. Ban quick a few paywalls sites, always put Wikipedia articles on top, prefer blogs than stackoverflow stuff…
There was little to no spam, though, but not much to look at either. Maybe it might be useful when searching for stuff that usually has high amount of confusing spam, but otherwise not really useful for me...
Kagi really shines when you are doing a standard search, though, which is what most people do most of the time.
And it seems like they also have a 1000 domain limit?
I understand the author's point of turning off their ad blocker "to get the non-expert browsing experience" but then they could make a different test with uBlock on for every query and see how it goes.
It's also a bit inconsistent to expect results for downloading videos mentioning yt-dlp while trying to emulate "the non-expert browsing experience"... Yt-dlp is a command-line Python utility. Talk about non-expert! Most people don't know that videos are files that can be downloaded; of those who do, most don't know about the command line or Python.
Yet when searching for "how to download youtube videos" the first result I get on Google is a link to a service called "savefrom.net", which appears to work well and does not seem to be a scam. This would qualify as "very good" in my book.
When searching for "how to download youtube videos from the command line" the first few results are about youtube-dl, including links to github and superuser. Granted they don't mention yt-dlp, but youtube-dl is a good start.
- https://msunduziassociation.online/perfect-online-videos/
- https://gssaction.org/program-all-in-one-media-solutions/
I would certainly put those in the "Terrible" category like the author.
It could also be related to targeting, like time zone, location, IP address, age group etc.
savefrom.net is the 5th result (2nd page underneath 5 youtube videos)
Edit: This is from the US. If i had to guess, these are regional differences. What country are you in?
Both seem to do the job of downloading a youtube link to mp4 for free.
Are you seriously suggesting that a website with the following "About us" with only a link to another YouTube video downloader is itself a good YouTube video downloader?
> Good Samaritan Support Action is to reawaken the Body of Christ to receiving the extravagant love of The Father, as well as our call to respond to this love by loving God with all of our hearts, souls, strengths, and minds. In order for people’s hearts to be linked to the heart of our Heavenly Father, we want to foster and facilitate the establishment of a culture of love in our churches and ministries.
Ideal? No. But it does the trick.
It seems pretty arbitrary to me to disable one of the key features - in this case personalization - of the software being evaluated.
Or is the evaluation not between "search engines" but rather "search engines without personalization"? If so, then this restriction does make sense. But that is not the evaluation that "normal users" are interested in.
It's the closest we can easily get to the 'average user experience'. Someone who has a long account/cookie history with Google has plausibly trained the site to return more relevant results through implicit user-curation of avoiding obvious-to-them SEO-spam on other queries.
If we posit that every user eventually trains Google to avoid SEO spam, then this begs the question of why Google(/Bing) don't eliminate the SEO spam in the first place.
Besides that, it's not obvious why search engine personalization should dramatically change the basic utility of search results. We should expect personalization to mostly address ambiguities: is 'the best way to set up tables' asking about furniture assembly/carpentry or SQL? None of the author's queries for this article supported such ambiguities, and besides that the results returned (see the final appendix) aren't[†] valid answers to a different interpretation of the question.
[†] -- I think I'd quibble about the 'adblock' question, since a reasonable person might still find an adblocker that works but participates in the 'acceptable ads program' to be sufficient.
You wouldn’t be really taking the average here though would you? You would be capture the experience someone might have if they were in incognito, using google for the very first time, or using google on another device for the every first time, but not the “average experience”.
Maybe it's the closest we can get (though I doubt it), but it definitely isn't close enough to tell us anything about the "average user experience".
The average user has been using google for years, without taking any steps to avoid personalization. An incognito session (on a browser / machine / network that is probably fingerprinted...) is pretty much the opposite of that typical usage pattern.
I recognize that just writing a blog post or comment on HN is not a research project so needs to do something quick, but I think it mostly invalidates the experiment. What would get closer would be to devise a few user personas and attempt to search and browse for awhile within those personas before trying the experiment. Or much better yet, put together a focus group comprised of real people within the personas you're interested in, and run the experiment using their real accounts.
> If we posit that every user eventually trains Google to avoid SEO spam
I don't think it's that, I think it's that every user trains it to return results more likely to improve the metric of "more likely to click one of the links", and I think that makes it more, not less, likely that they see what most of us here consider to be spam.
But I don't know! Maybe that's not what this experimental setup would show. But it would be a lot more enlightening than a setup using a fresh incognito window, which reflects the usage pattern of a proportion of search queries that is a tiny rounding error above zero.
If this doesn't describe most people you know, you're in a very small bubble. (I'm somewhat in that bubble too, but I still have lots of family and friends who use the internet the normal way.)
In this thread we can see people both using incognito tabs seeing different results, it will only become worse to compare if they are using personalized results.
(https://www.linkedin.com/pulse/how-we-increased-search-accur....)
But if clicking a site meant I would be under attack, that really increases the stakes, I start to care strongly about the absence of bad sites, not just the existence of a good one.
Other than that, people need to be trained to not download programs from websites in general. I think this has gotten better over time? This is just a human mistake. Maybe Google could suppress sites that link to executables. It must, right?
The takeaway I got from the article is everyone can make their own test, as opposed to relying on other people's sentiments and memes about X is bad or Y is good.
Trying to emulate a non-expert experience without workarounds is not the common usage pattern since everyone familiar with their favorite tools have ways to get more value out of them, but this article presents a way of constructing an experiment (this is why I chose these queries, this is how I ranked scams, etc.), and I think people should follow this same spirit to evaluate if they are stuck in a local optimum with their current choice of tools.
Just give me a website where I can plug in the DL link and download it to my hard drive. I don't care what package they are using (I don't worry about malware like I did in the 90s). 99.999% of people are not programming tinkerers.
Just makes me realize how subjective search results are. All of their "Great" results are my "Terrible" results.
But again, I guess that's why search is so hard is because I have to parse that intent from 3 words.
It's pretty simple to run these download tools as website, but it's expensive in terms of bandwidth and tends to attract legal attention. So a lot of websites go up supporting it, but even if they were started with good intentions, they will virtually all eventually add intrusive ads or other types of monetization just to break even. So there's never going to be a reliable website for it. If you're lucky, a search engine will send you to one that's working okay right now, but even odds you'll be fighting through a dozen malware nests.
Meanwhile, yt-dlp just works every time, with only an occasional pip upgrade to keep it up to date.
Like, sure, I have the impression that search got worse over the last years, but .. has it really? How could you tell?
And, honestly, this should be a verifiable claim; you can just try the top N search terms from Google trends or whatever and see how they perform. It should be easy to make a benchmark, and yet no one (who complains about this issue) ever bothers to make one.
Dan at least started to provide actual evidence and criteria by which he would score results, but even he only looked at 5 examples. Which really is a small sample size to make any general claims.
So I am left to wonder why there are so many posts about the sentiment that search got worse without anyone ever verifying that claim.
I encounter claims that "protobuf is faster than json" pretty regularly but it seems like nobody has actually benchmarked this. Typical protobuf decoder benchmarks say that protobuf decodes ~5x slower than json, and I don't think it's ~5x smaller for the same document, but I'm also not dedicating my weekend to convincing other people about this.
So what people are actually saying is, a typical protobuf implementation decodes faster than a typical JSON implementation for a typical serialized object -- and that's true in my experience.
Tying this back into the thread topic of search engine results, I googled "protobuf json benchmark" and the first result is this Golang benchmark which seems relevant. https://shijuvar.medium.com/benchmarking-protocol-buffers-js... Results for specific languages like "rust protobuf json benchmark" also look nice and relevant, but I'm not gonna click on all these links to verify.
In my experience programming searches tend to get much better results than other types of searches, so I think the article's claim still holds.
If he was looking at relevance, yours would be a solid point, but since most of the emphasis is on harm, a smaller sample works. Like "we found used needles in 3 out of 5 playgrounds" doesn't typically garner requests for p-values and error bars.
That's bad for google though! Their model is very much predicated on the web having a lot of signal that they can find within the noise. But if it just ... doesn't actually have much signal, then what?
Imagine Google were like a water filter you install on your kitchen faucet to filter out unwanted chemicals from your drinking water. If as the years progress your municipal tap water starts to contain a higher baseline of unwanted chemicals, and as a result the filter begins to let through more chemicals than it did before, you'd consider your filter pretty cruddy for its use case. At the bare minimum you'd call it outdated. That is what is happening to Google search
I'd also argue that finding relevant results among a sea of irrelevant results is the primary function of a search engine. This was as true in 1998 as it is today. In fact, it was Google's "killer feature", unlike Altavista and the likes it showed you far more relevant results.
Not saying or even suggesting that's happening, but the logic isn't airtight
These days the results are apt to be the correct topic, but instead optimized for some other metric than what the user wants. For example downloading malware or showing as many crypto ads as possible.
I don't expect every search engine to have the same scam results. Scammers target individual search engines with particular methodologies. Google does a lot of work to prevent crap on their engines, the issue is the scammers in total do far more.
Or maybe we're saying essentially the same thing, but you think search engines should be doing that curation. But that was never my conception of what search engines are for.
I'll give Google credit: I haven't seen gitmemory or SO clones in a while. It took a few years but they seem to have dealt with them.
There still is a question about when it got bad--I think Dan mentions 2016 as a point of comparison, and there were plenty of scams back then, so you might wonder whether the days when a query wouldn't return many scams.
If you go back far enough, then there wasn't the same kind of SEO, and Internet scams were much smaller/less organized, but that's a long time ago.
Yes, and he makes the point well. It also means if you are part of the 0.49% of people who use Firefox on Android, he isn't talking about your experience. I find Firefox mobile remaining at 0.49% utterly inexplicable, which I guess just goes to show how out of touch with the mainstream I (and I assume most other people here) are.
It's not just ad blockers. My first attempt at a tyre width query got relevant results, mostly because "tyre grip" looked so bad as a search term so I used "traction" instead. In the mean time, friends of my age (60's) can't get an internet search for public toilets to return results they can understand. When I try to help them, their eyes glaze over in a short while and they wave me away in frustration. These mind games with google hold no interest for them.
I am regularly bitten with one thing he mentions: finding old results is hard, and getting harder. It makes it really hard to find historical trends ("am I wrong about what it was like back then?") really difficult.
1) The step where you evaluate "how they perform" is necessarily subjective.
2) you could design a study and recruit participants but that isn't something a blogger is going to do.
3) He does link to polls where people agree with the idea the result have gotten worse. Yeah, there are sampling problems with a poll, but its better than nothing.
In this case especially, the writer is answering the question: "Whose results are best according to my tastes?"
Find a query of interest, see for yourself (and take a snapshot of the present state for posterity).
The api enables more powerful queries, https://web.archive.org/cdx/search/cdx?url=google.co.jp*&pag...
Also try other search engines and languages.
US NIST, in their annual TREC evaluation of search systems in the scientific/academic world, use sets of 25 or 50 queries (confusingly called "topics" in the jargon).
For each, a mandated data collection is searched by retired intelligence analysts to find (almost) all relevant result, which are represented by document ID in general search and by a regular expression that matches the relevant answer for question answering (when that was evaluated, 1998-2006).
Such an approach is expensive but has the advantage of being reusable.
All that matters is the overwhelming sentiment that search has gotten worse, not the same fucking spreadsheet that got us here in the first place!
I suspect it has gotten worse, so posts complaining about it resonate. But, it is not really a huge problem, and anyway it isn’t as if there’s much I can do about it, so I’m not going to bother collecting statistically valid data.
I think this is generally true about a lot of things. We should be OK with admitting that we aren’t all that data-driven and lots of our beliefs are based on anecdotes bouncing around in conversations. Lots of things are not really very important. And IMO we should better signal that our preferences and opinions aren’t facts; far too many people mix up the two from what I’ve seen.
I can't speak for anybody else, just trying to find stuff online, not writing a treatise about it or writing my own engine to outcompete Google. It's been asked many times here over the years and the answer was always explanations, never solutions.
Shittification does not happen overnight, but along many years. It started with Google deciding that some search terms weren't so popular: "did you mean...?" (forcing a second click to do what you intended to do in the first place) and went downhill when qualifiers to override that crap got ignored.
For me enough was enough when I realized that a simple query with three words, chosen carefully to point to the desired page, gave thousands of results, none of them relevant. YMMV.
That is, in the early days, Google used to highlight that "search position couldn't be gamed/bought" as one of their primary differentiators, ads were clearly displayed with a distinct yellow background, and there weren't that many ads. Nowadays, when I do any remotely commercial search the entire first page and a half at least on mobile is ads, and the only thing that differentiates ads from organic results is a tiny piece of "Sponsored" text.
Otherwise, yeah, maybe search didn't degrade but the internet got more spammy. Or maybe users just got wiser and can see through the smoke screen better. Who knows...
Doesn't change the fact that today one has to know how to filter through pages of generic results made by low effort content farms. Results that are of dubious validity, which at best simply waste your time. Or through clones of other websites (i.e. Stackoverflow clones).
Search engines can choose to help with that (kagi certainly puts in the effort and I love it for that), or they can ignore the problem and milk you for ad clicks.
Anecdotal evidence is good enough for me.
I still remember myself having to really often go to page 3 and more of google searches to find things even really early on.
I think it has never been good, got a bit better before SEO farms took all the gain out. That's my feeling with nothing to back it.
For example, let's say I search for "Gaza"; on one extreme end some engines might only focus on recent events, whereas others may ignore recent events and includes only general information. Is one higher "quality" than the other? Not really – it depends what you're looking for innit?
All you can really do is make a subjective list of things you find important and rate things according to that, and this is basically just the same thing as an anecdotal account but with extra steps.
Yes it has and for a certain class of queries it's not even open for debate, because Google themselves have stated they deliberately made it worse. And they really did, it's very noticeable.
This class of queries is for anything related to any perspective deemed "non authoritative". Try to find information that contradicts the US Government on medical questions, for example, and even when you know what page you're looking for you won't be able to find it except via the most specific forcing e.g. exact quoted substrings.
Likewise, try finding stories that are mostly covered by Breitbart on Google and you won't be able to. They suppress conservative news sites to stop them ranking.
15 years ago Google wasn't doing that. It would usually return what you were looking for regardless of topic. There are now many topics - which specifically is a secret - on which the result quality is deliberately trashed because they'd prefer to show you the wrong results in an attempt to change your mind about something, than the results you actually asked for.
There are particular search parameters that DDG changed the behavior of, including exclusion and double quoting, which are now, according to even their own docs, more a hint of the direction results should go rather than any explicit/literal command (ime these virtually never work, which was a motivation for documenting failures, and they actually removed them from their docs temporarily at one point earlier this year).
https://static.googleusercontent.com/media/guidelines.raterh...
It talks about figuring out a query’s meaning(s), judging the user’s intent (were they looking for some specific answer, etc.), evaluating the “quality” of a website, rating the site’s usefulness in relation to the query’s meaning/intent, etc.
All this is to say, it’s not that search companies don’t do exactly what the author did here, it’s just that they have different standards than the author. And I’d venture their standards match their users’ better than the author’s, but maybe not or not forever, anyway.
My hope is as LLM’s improve, they can be more discriminating about the results returned.
I didn’t say they would :)
In fact, I can’t figure out how your comment relates to mine. Are you claiming that Google doesn’t factor blog spamminess into its evaluation of search results? If so, that’s quickly put to bed by the document I linked, pretty much section 4.6. Excerpt:
> Creating an abundance of content with little effort or originality with no editing or manual curation is often the defining attribute of spammy websites.
You could claim that they fail to capture some essential quality of “blog spamitude” or that they don’t weight it heavily enough in their eval but to say they just, like, don’t know about blogspam over there, is pretty far fetched IMO.
While I obviously don't know it may be related to how Google believes a "normal" person search. I have come to view Google as a product search engine/price comparison site, that's what it's great at. Google can find you the most relevant products for any purchase you may consider, so maybe that's what Google has optimized for. The majority of my searches are related to IT, programming, software and computers in general, but what does "normal" people search for. They search for products, news, opening hours for a store, Google is pretty decent at that, but the money is in the "go buy something". The ads on a product search on Google is always way more accurate than the actual search result.
I think Google has optimized for selling products.
Different people have varying expectations as to what they want to find with the same query. I'd definitely want yt-dlp in favor of some website.
But can the search engine mind-read by assuming Windows users don’t want to use a command-line utility?
In the Youtube Downloader search, NortonSafeWeb is nowhere to be found. I get a couple of legit downloader websites, and some articles from reputable tech newspapers on how to use them or command line tools.
In the Adblock search, ublock Origin is #3, followed by some blogs about ad blocking ethics debates and the bullshit Google has been pulling recently.
In the wider tires grip search, #3 is a physics blog that dives deep into the topic.
In the transistors search, the first reddit link directly answers the question in very similar wording to the hypothetical correct answer spelled out in the rubric. 4/5 of the reddit results are on the correct topic, followed by two SuperUser questinos also on the correct topic, then some linus tech tips and toms hardware articles, also on the correct topic. No Quora questions.
In the vancouver winter snow search, the first several results are from local news papers talking about the anticipated effects of el nino on snowfall, and then a couple of high-quality blogs and weather sites.
Really wondering how Dan got such bad results.
------
Aside from that, the way that the author expects all the results to return the same kind of thing is just... weird? Like, that's not how search engines are supposed to work. A search that gives you 10 links to fundamentally the same thing is a bad search. Search results should cover a breadth of reasonable guesses for what you should be looking for given a query. If you search for "download firefox", and you scroll past the first 5 download links, then you're probably not actually looking for a download link and a blog post about firefox is not "irrelevant" and shouldn't be points against.
This opinion is even borne out in search engine quality metrics that have been industry-standard for decades, like mean reciprocal rank and distributed cumulative gain. What matters is how far you have to scroll to get to a good result, not what proportion of the first N results are good.
Ctrl+F on the page for "System prompt" doesn't show any hits. Given how important those are for ChatGPT (another thought - was the author testing GPT3.5 or 4?) I'm not sure how much weight to put into the ChatGPT results either.
Not sure how much I can take away from this comparison.
Getting any useful data from GPT-4 about anything even remotely “illegal” is a waste of time.
It suggested '4K Video Downloader', 'YTD Video Downloader', 'JDownloader' and 'Clipgrab' at first and when I asked for cli tools it came with 'youtube-dl', 'yt-dlp', and 'ffmpeg'
Those seem pretty reasonable results to me but I'll readily admit I don't know (yet) if 'most users' would ask these follow-up questions.
Mistral showed that their medium model is far better (yet not good), and the same prompt as in the article gives only one instead of 3 paragraphs of rambling about copyright, and then lists 3 categories of options with examples for each (not good, because ytdl is not one of those listed).
Funnily enough, both mistral and GPT4 apologize profoundly and almost with the same wording when asked "Why did you not mention the very popular, free and open source "youtube-dl" software?" and then mention how/where to get it and how to use it.
Likely because they were optimized for general population, which would not have a use for command line python utility.
For the transistor query, it’s a very "googly" way of writing a query, when I saw the results I instantly felt like rewriting it and the first try gave much better results with "Why keep cpu transistors getting smaller?". Caveat that the results look better and more topical, I don’t know what a good answer would be, also why I didn’t evaluate the tires or Vancouver weather (I tried a local search for my cities weather, and while the first result was unreleated, the 2nd was okay)
edit: This whole thread made me finally create a file for documenting bad searches on Kagi. The issue for me is usually that they drop very important search terms from the query and give me unrelated results. But switching to verbatim or "forced terms" also prevents any kind of error correction of the search. This used to be one of my main annoyances with DDG back then, and Kagi did not have that issue during the early days.
A specific example: for "ad blocker" the first result was some paid ad blocker and ublock was down the page below the fold.
It's already performing better on a (n=1) test I tried.
"Talos Principle 2". (Video game sequel) Previously (~5 days ago), Google returned various screenshots etc from the game `The Talos Principle 2`. Kagi returned mostly results from `The Talos Principle (1)`. Now the latest Kagi results are a mix, mostly from 2. So, it does look like it fixed this query.
Try https://search.metaphor.systems - it's fully neural embeddings-based search. No keywords, only an embedding of what the actual content of a webpage is.
So in the mentioned example of searching for Youtube downloaders, with Metaphor you'll get only Youtube downloaders (https://search.metaphor.systems/search?q=This%20is%20the%20b...)
Full disclosure - I work there :p
What prevents websites from gaming their embedding? Switching to a similarity search doesn't prevent the results from being gamed.
edit: The results are also from my quick QA not that great. Searching for "what is the best mouse to buy" leads to links to buy random mice versus review summaries or online discussions on mice. One of the recommended queries of "Here is a great fun concert in San Francisco" leads to some really bizarre results in non-English languages that have nothing to do with either SF or concerts.
edit2: Also, Google has been using LLMs part of their search since at least 2018 so definitely not just keyword matching there.
For your search - I would recommend turning autoprompt off and searching something like "Here is a great summary of the best computer mice to use:".
Our embeddings model is trained on how links are talked about on the Internet, if that helps with querying. So you have to query like how someone would refer to a link before sharing it
So it's not high quality web pages but web pages that people talk about a lot which is expected since no one has an oracle that says what high quality is. The embeddings are merely a proxy and generalization for "how links are talked about on the Internet." That can be gamed at scale just like every other signal any popular search engine has been based off of.
Definitely excited to see how it holds up to daily use.
So far it gave me exactly what I wanted at the top for all of my test queries that were well formed.
As for asking “ignorant” questions both your service and the goog failed where phind gave me an actionable starting point (after a prodding follow up question: https://www.phind.com/search?cache=hmul4znpn7y4ei6qa64fosmc )
“max-height like css property for top and left”
Unsure if this sort of thing is even a goal of your project, but you won over a new user.
Wish you and your team all the best.
For paywalls/login - we play pretty straight, always obey robots.txt, etc.
Auto-prompted to: "Here's a helpful website for downloading YouTube videos:"
Also, this result is horrible:
“What does it mean if someone is not covered in nfl football?”
I haven't installed the downloader app, so I'm not sure if it lets me download youtube videos for free.
The second result "ytder.com" is a redirect to "https://poperblocker.com/edge/" which seems to be a browser extension for Microsoft Edge that protects the user from the Holy See. I'm not using Edge and I'm trying to download a Youtube video.
The third result download-video.net says that it can download videos from a list of sites. Youtube is not in the list, but let's try anyway. If you put "https://www.youtube.com/watch?v=IkYVmtgxebU" into the text box and click "download" you get "500 SyntaxError: Unexpected token '<', ""
At this point I gave up, but please let me know if any of the results work.
I clicked into the top 5 results, none of them were real youtube downloaders that worked, so I clicked the next 5 results, then I finally got one single (really slow) downloader that worked. 1 out of 10 top results
That is not to say that there aren't valid criticisms of and shortcomings in ChatGPT 4 - just that it's not useful to say ChatGPT when it's referring to 3.5
Also I would not normally write ChatGPT queries the same as I write them for search engines but for the sake of comparison, I'll use their queries verbatim except where my custom instructions affect the context too much.
> download youtube videos
https://chat.openai.com/share/3e18e4f0-5527-4479-8a2f-ef17bd...
I got - good results. They got - "Very bad results (fails to return any kind of useful result) ChatGPT: basically refuses to answer the question, although you can probably prompt engineer your way to an answer if you don't just naively ask the question you want answered".
> [What] ad blocker [can I use?]
https://chat.openai.com/share/e1985d7a-c89f-4b5e-bb59-70bd11...
Looks good to me
> download firefox
https://chat.openai.com/share/3a62e5ae-8dbd-4179-8eb0-cc38ee...
Also good
> Why do wider tires have better grip?
> [Provide links to scientific sites that describe] why wider tires have better grip?
https://chat.openai.com/share/8cbcd1dc-b23f-41f3-83ad-f43f3d...
Honestly, I have no idea if this is a good answer or not. But I don't use ChatGPT for answers that I don't have confidence that I can determine its veracity; if I needed to know this with certainty, I'd use ChatGPT as a jumping off point for my own research.
> Why do they keep making cpu transistors smaller?
> [Provide links to scientific sites that describe] why do they keep making cpu transistors smaller?
https://chat.openai.com/share/dbb97ac0-840c-402c-a917-657af6...
> vancouver snow forecast winter 2023
> Environment Canada winter 2023
https://chat.openai.com/share/aab017d7-f86b-49c9-b5c0-86a0b1...
I don't know if almanac.com is any good but giving it the specific "Environment Canada winter 2023" query gave the expected very good result.
I think ChatGPT 4 generally provided very good results for the test queries, if you tailor the queries just slightly for the format
Does everyone or even most use ChatGPT 4? The most used version is -of course- by far the most relevant.
But I suppose what I really want is for everyone who includes ChatGPT in comparisons to explicitly say which version they are using (and, if they are using 3.5 in their comparison I hope they at least try 4 first) and definitely not just say "ChatGPT" when they only mean 3.5. The difference really is that stark.
A better benchmark would have had two entries for ChatGPT, showing both 3.5 and 4 results
And yet to other people it starts rambling about how that’s wrong and you shouldn’t do it and doesn’t give a usable answer.
https://news.ycombinator.com/item?id=38822040
I boggles the mind the extent to which people salivate over a system that cannot decide between a correct straight answer, something wrong but plausible, something wrong and impossible, or outright refusing to answer.
For what it’s worth, I do have access to v4 and it did give me an answer right now. But since I also know even v4 can give you wildly different answers to the same question even if you ask them one right after another, that doesn’t prove it either way.
Is that true? Do most people want to install a command line tool to download youtube videos?
After that there is a popup asking me if I want to continue in the browser or download their app. If I click download, it downloads a file called download_helper_2.3.27.apk.
Instead of downloading their app, if I paste a YouTube link, it tells me I can wait or download their APK to skip waiting. The download link downloads an older version called download_helper_2.3.19.apk.
When I do the process again, instead of the older APK link it gives me a Chrome extension link. But if you look at the instructions you see that it's not a Chrome extension, but a minified userscript. And it has `@include https://*` so it can basically run on any website regardless of clicking on an extension icon like regular browser extensions.
If I try to ignore all the distractions and wait for the download link, I can click it and it downloads the MP4 file. But it also opens a popunder with the domain https://refpamjeql.top/.
Not the best experience, and seems like a high risk of getting malware, but it does get an MP4 file at some point.
Edit: the paid subscription payment flow says I'm actually buying "Televzr Premium Max Subscription for
1 Month_mp
Televzr helps get wireless access to the media library on the computer from the mobile phone"
So it purports to be something unrelated to downloading youtube videos. I didn't pay 1400 yen for it, so I won't get to find out if it helps me download youtube videos.
Having said that, I’m usually still able to find what I’m looking for, if I know that it likely exists, and know the keywords to use to find it. But it’s much harder nowadays for sure.
I also wonder if google just stopped existing, would the web heal over time?
I don't buy the signal to noise argument. For example, whenever i get on youtube and get fed some content, i can immediately tell if it's had AI involved anywhere, and thumbs it down. I won't recommend it, i've called people out for linking such tripe to me (or others).
Hear me out - google got bad about 11 years ago when the dorking stopped being effective, right around the time of the spotlight search results and the sponsored junk taking the top results. Around this time, various agencies (news, etc) started gaming the SEO to respond to any remotely related search with whatever the news was currently. Google chose not to "fix" this, because we're not the customer. DDG was better for a few years for real results, too, but that has gone downhill as well.
The current zeitgeist uses stuff like tiktok and facebook for "web searches" - "food trucks near Austin, TX" or so. No one really uses web search like people on this site do, and google couldn't care less if we don't like the search results.
I have a hard time believing an organization like Google doesn't have the resources to provide a search engine that's just as usable as what they had 6 years ago (around the time I feel like the decay really set in). Seems a lot more likely that it's just more profitable to serve up garbage sponsored content.
I’d rather see a world with numerous paid/subscription search engines, that are motivated to do nothing but return search results well. I expect you would see some of the SEO crap getting solved.
On the other hand, they can achieve virtually the same outcome while keeping plausible deniability by just not doing anything that would downrank sites with ads (of which a significant chunk is likely to be Google's).
Spam sites often include ads.
Therefore the safest option is to never openly discuss it or intentionally do it and instead use other means to achieve the objective (don't intentionally rank spam higher, just defund/cancel any projects that would make it rank lower).
Yup, and I think we've seen how careless and thoughtless Google is, as an organization, with internal comms in the Epic case. It would be shocking if it hadn't been discovered in that, or a prior, lawsuit.
So it's inevitably going to become crap.
But all in all, a very good article
EDIT: here's a search for tire (I don't know anything about tire, so maybe there's much better links out there, but this is pretty much what I was expecting. Not an ad or SEO in sight.) https://www.perplexity.ai/search/tire-3iuI9T6BQUSvu2tAhgsRmA...
For example, the first query is "download YouTube videos", for which Google is ranked "terrible" for not showing you a command line open source program. But the literal first result is an ad supported site where I can paste in a YouTube link and download it right from the browser. That seems like exactly what most people would want or to the CLI tool the author is searching for. The author seemed to be looking for sites without ads as what they wanted to see in search results more than search relevance.
Search is a very gamed system with a lot of SEO spam type results, but I think a much better analysis could be done for more meaningful results. Also I recreated some of the searches and got very different results (including ublock origin in the top three responses). Again, a more scientific ranking system could help uncover better data on searches.
> Some youtube downloader site. Has lots of assurances that the website and the tool are safe because they've been checked by "Norton SafeWeb". Interacting with the site at all prompts you to install a browser extension and enable notifications. Trying to download any video gives you a full page pop-over for extension installation for something called CyberShield. There appears to be no way to dismiss the popover without clicking on something to try to install it. After going through the links but then choosing not to install CyberShield, no video downloads. Googling "cybershield chrome extension" returns a knowledge card with "Cyber Shield is a browser extension that claims to be a popup blocker but instead displays advertisements in the browser. When installed, this extension will open new tabs in the browser that display advertisements trying to sell software, push fake software updates, and tech support scams.", so CyberShield appears to be badware.
It's a service that is quasi illegal and explicitly breaks the YouTube terms of service. I think the search engine did a good job surfacing what was searched for, there just aren't going to be any free online YouTube downloaders without advertising.
It seems like the author wants to search to read their kind without specifying what kind of YouTube downloaded they want
> Download youtube videos
> Ideally, the top hit would be yt-dlp or a thin, graphical, wrapper around yt-dlp. Links to youtube-dl or other less frequently updated projects would also be ok.
That's not what a random person expects. yt-dlp or youtube-dl have no meaning to a normie. The first result is an online downloader and that's what an average person is after. I checked the first result in Kagi and it's a valid youtube downloader.
If you're after a commandline tool, ask for it: "commandline tool download youtube videos" gives youtube-dl as the top result with valid options afterwards: https://kagi.com/search?q=commandline+tool+Download+youtube+...
"Ad blocker" seems to ignore other options exist. Yes, ublock would be preferable for most, but ABP is not "very bad". Kagi mentions ABP at position 1 and ublock at position 8: https://kagi.com/search?q=Ad+blocker&r=au&sh=4VHApDrTEfuxMOt... (But for a query like that, I'd be happy with a wikipedia article about adblockers, because why not?)
I'm not disagreeing that results have been getting worse for years, but... this is a really bad scoring system. It feels like that one very new person jumping on SO posting something like "syntax error: if 1 {" - what are you even asking for? (To be honest, the search engines could also give you the equivalent of "this is a very vague, would you like to specify what you're actually after? here are some suggestions: ...", but that's beyond the scope here.) The search returning not the exact thing you want to see for a super generic query, but returning a valid answer to a question is not "very bad".
First one is https://wave.video/convert/youtube-to-mp4-355 It downloaded the first video just fine.
Like, for me at least, I already know yt-dlp exists. When I search "youtube downloader", it's exactly because I want an online-website page to download youtube videos.
I've been using Kagi's FastGPT [0] now for these searches, it basically removes all the bullshit and gives verifiable sources for any answers.
Added: Interesting. Apparently it's allowed to edit the SERPS there. Which implies that I'm out, but (well, because) I've got a feeling which kind of Internet Entrepreneurs this factoid will appeal to
For instance, I wrote an R ggplot2 package called "fedplot" (following the convention of calling the package for the figure style it replicates, as in "bbplot" for BBC-style charts).
Try searching for it on Google: "github" "fedplot" doesn't get you anywhere. Meanwhile, every other search engine gives you exactly what you want if you just type "fedplot". I even tried to add the relevant websites through google's suggested tools, and nothing happened :|
Who needs to know anything about government owned land anyway?
Qwant: Result 1
Bing: Result 1
Google: Result 2
Marginalia: Zero results
ChatGPT 3.5: Some Federal Reserve dot plot nonsense and no useful results.
Meanwhile, both Bing and Qwant give me exactly what I want
For both the tire question and with respect to a youtube dowloader, the first results were on the nose with the addition of site:edu on Google.
Why this is needed and whether a noncommercial, information rich web portal should exist are questions for another thread.
Better answer: learn the differential equations in this book:
https://ftp.idu.ac.id/wp-content/uploads/ebook/tdg/TERRAMECH...
Brave search API is obscenely overpriced. I hope someone is working on Search because Google has become a singularly garbage company. Propping up DEI is sinful enough but just failing to compete is lame. /shrug
I'm probably in a very small group who have the entirety of English wikipedia (without images) on my Android (via Kiwix), and I just search that. 99% of the time that's all I need.
the only exceptions are super current things like weather (Windy), or travel (Navan work travel system gives me enough to just go direct to airlines, hotels, etc), and local (OSM via Organic Maps).
I've almost completely degoogled (not intentionally, but driven gradually by Google becoming crappy incrementally), but didn't really find a single generic replacement as much as I found far better single purpose tools.
I'm reminded of that Craigslist image showing how many startups were each competing against specific parts of Craigslist https://cbi-blog.s3.amazonaws.com/blog/wp-content/uploads/20... , and this is what it feels like is happening to Google.. they're being beaten in specific areas, but at the same time spam and crap is diluting their core product.
Not saying the current landscape doesn't suck with ads everywhere and incentives to not give exactly relevant results at times, but I think google is pretty good still.
From what I understand, it aggregates results from multiple sources rather than having their own indexer.
The results aren’t really any better, but the lack of ads and videos in the results makes for a cleaner experience.
I also haven’t yet taken advantage of the extra features to block certain websites from results.
Personally, I pay the $5 mostly in an attempt to support another competitor in the space.
Start using bangs, lenses and customized results ASAP, that makes a big difference.
Which is a similar phenomenon to search. If you have sufficient tech skills there's a whole world of freely available software out there to complete your task.
If you're not then you are at the mercy of a range of commercial offerings (some built on the free software) that range from arguably scams to outright scams.
When I search for "Steve Jobs" on Marginalia, I got blogs about his speech in 2011 and some mailing list from 2007.
When I search for my own name I get nothing. In Google it's just me.
It's cool that one person built all this of course but... that's not a good search result compared to Google?
Maybe I miss something, maybe I use it wrong
I expect wikipedia article on Jobs as a baseline.
It's amazing what you did, it's just not a Google killer? or at least I don't see it
In general a lot of the complaints seem to be "I'm not getting what I expect from Google". Well... yeah. That's the point. If someone wants the same results as Google, they should arguably use Google.
Interestingly, if you add "before:2001-01-01" to the query, the paper that Brin and Page referenced shows up as the third result.
That this query now ranks phones you can buy higher than information about phones makes sense, since the web is much bigger these days and cell phones are much more widely accessible than they were back then.
> Although Google doesn't publicly provide the ability to see what was historically returned for queries, many people remember when straightforward queries generally returned good results.
See above. Sort of.
---
I wish Dan spent more time talking about Kagi. I, too, have found it terrible for searching for things to buy and some images but excellent otherwise.
"Two of the top three hits are how to install the extension and the rest of the top hits are how to remove this badware. Many of the removal links are themselves scams that install other badware."
I think that this very sentence shows the author's bias, because I feel that Google's search results are not just great, but better than what it was 10 years ago.
Unlike the author, I think that building a better search engine than Google is possible. But it's going to be rather expensive. And the only proven way to monetize it is selling ads. Which will degrade the quality of the search results fast. For potential investors, there are probably many better ways to invest money then by building a search engine.
This lets us with only one viable alternative: build it in the open like Wikipedia and source donations from people and from Google competitors like Amazon or Apple.
Finally some one said it. We are unnecessarily harsh on hallucinations. LLM’s don’t intentionally ‘lie’. To say this is a wrongful anthropomorphism.
We're not unnecessarily harsh on hallucinations, it's absolutely necessary because of how effective LLMs are at convincing people that because they can generate language, they are capable of sentient thought, self-awareness and reason. Acting as if humans and LLMs are basically equally trustworthy, or worse, that LLMs are more trustworthy, is dangerous. If we accept this as axiomatic, shit will break and people will die.
I don’t second guess Pythons math results. If the result is wrong, that’s my fault for coding it wrong, never Pythons for hallucinating
And changing the query to "ad blocker" like Google suggested raised ublock origin way up in the results
Also, the article tested Mwmbl as well, not mentioned in the title here.
If the topic has ever come up the discussion and links are likely to be more relevant and better than your avg. wiki article
As I understand it, this is because tyres are still somewhat of a mystery, and anyone outside of a laboratory really doesn’t know shit. The best explanation I can think of is due to tyre load sensitivity. The friction coefficient of rubber decreases with normal force (E.g, a heavily loaded tyre has a lower friction coefficient), which is a pretty well accepted fact, this is one of the methods engineers will use to tune the handling of cars. This means a wider tyre has a lower force per unit area of the contact patch, which means it’ll have a higher friction coefficient.
Now that sounds plausible to me, but that’s just my best guess explanation.
gives good tyre advice (obviously not car tyres, but info is there)
1: my reading is that this is a sarcastic denomination for someone that is supposed to be an innovation thought leader but actually is just defending the broken search landscape status quo.
Maybe LLMs will help, but I can’t shake the nagging feeling that the situation will simply get worse with LLMs, not better, due to hallucinations and the apparent “gullibility” of LLMs: I would not be surprised if SEOing an LLM turns out to be easier than SEOing Google.
Searx and Yandex.
Specifically… if I need something even slightly “gray”, Yandex is the only option anymore. Torrent search on google et al is just awful.
why
Without business, spam would disappear.
So if you remove the labor you remove the spam.
So the best spam filter is UBI.
Ublock origin in the very top result for ios device is simply a bad search result page. Maybe fourth position is tolerable, after three different working ones. Maybe it should be lower, I doubt myself, if my point of view is too elitist.
Yt-dlp is subject to all sorts of takedown requests in different jurisdictions.
I feel like today's digital spaces don't have as strong a grip on the minds of people - I think folks started rediscovering the value of genunine human interaction and hobbies that do not involve a computer screen.
For example, I haven't seen the equivalent of 2000s-2010s Facebook addicts or (WoW addicts in the gaming space) to such an extent, with parasocial media, such as TikTok or Youtube or Twitch, having replaced social media, and social video gaming such as MMOs having lost a lot of popularity.
It's most frustrating with phone numbers. I picked up the habit of searching the random numbers that called me, to try and find out if they were possibly important. I used to get a bunch of spam sites that clearly existed to profit off me making those searches.
Both Google and DDG have removed those spam sites, even though they were useful at times. Google will tell me the number is in some random PDF that contains a few of the digits, then no other results. DDG will say the top result is my local police department, something that freaked me out the first few times.
Query: “I’m coming out of my cage…”
Result (Ad): “You’ll be doing just fine with these amazing year-end closeout prices at Al’s Discount Car Barn. Gotta come down—you’ll want it all!”
When searching for results from my country in DDG (picking the country in the drop-down below the search box) still returned results from the USA or other countries. Even when searching stuff in the local language. Maybe they tried to fix that because it really sucked, so much I never used it again for searching into local websites.
However, more generally, I've personally found that DDG (and maybe Bing's then?) localised results are just really bad, and have been for the multiple years I've been using DDG and it's had this feature: I'm in New Zealand, and enabling localised / region-based search still often provides results to pages with TLDs like "co.uk", ".ca" and ".pl" (these latter are really common for content-generated spam in my experience), which I just can't understand...
Unfortunately, I have found that Google's results are usually a lot better in terms of being "location-aware" than DDG, at least when that's what you want...
You can report them: https://ised-isde.canada.ca/site/canada-anti-spam-legislatio...
Rather than wanting Kagi to take the place of DuckDuckGo, it would would be better if Kagi could take users from Google, and then when ready, drop Google as a search provider.
It's been available for ages. We used it to power the company internal search for a large enterprise I worked at 17 or 18 years ago.
https://help.kagi.com/kagi/search-details/search-sources.htm...
Perhaps the incorrect thing is not your internet search results, but actually your phone carrier for lying to you and telling you that a caller has a local number?
The top result being my local police department because it shares the same area code and has maybe one other number in common is clearly a bad result. It does this even if the phone carrier isn't lying to me and the caller does have a local number, like the increasingly common political spam calls.
I’ve also noticed a significant increase in attempts to stuff news into regular search results. I really do not appreciate being force-fed mental health poison. I don’t need it ever, but I especially don’t need it when I’m searching for some specific technical thing and then get emotionally sabotaged by some clickbait headline because … why? Some bullshit KPI? Why are tech companies so obsessed with pushing news into every orifice?
From what I can tell this is an issue with the Bing API that DDG uses that the DDG folks have been unable to resolve. I've tried many identical queries between DDG and Bing and while Bing does occasionally return incorrect local results, the completely irrelevant local results that appear on almost every DDG search do not seem to happen with Bing itself.
From what I understand, DDG is aware of the issue. I don't know why it isn't more of a priority.
Are you by any chance using a VPN while using Brave Search? (ProtonVPN?)
Thanks for your help, we're working on ways to reduce the number of captchas shown to VPN users and your feedback is very useful.
#1 "Gordon ramsey" (misspelled "Gordon Ramsay"). Marginalia shows "The Life I Imagine: are my cheeks red?". Kagi corrects to Gordon Ramsay and shows relevant results.
#2 "Ukraine war". Marginalia shows an article about the Russian Orthodox church and a Substack post about the war. Kagi shows Wikipedia, Al Jazeera, etc up-to-date summaries about the war.
#3 "Dildo". Top post on Marginalia is "Students for Concealed Carry Embraces UT Dildos | Students for Concealed Carry". Top posts on Kagi are Wikipedia (read) and Amazon (buy).
> How is Marginalia, a search engine built by a single person, so good?
Because it's not good?
[0]: https://physics.stackexchange.com/questions/29903/why-do-peo...
Edit: I did just realize that I have StackExchange customized to be up-ranked. So that probably helps. But yeah, I guess this is why I usually get good results, which is something that generally still fails with Google for me.
That's bad if you're looking for a simple answer or basic fact, and good if you're looking for a few hours of reading.
Using an adblocker is not expert anything.
That you've defined your own opinion for what some of the results should be blows the thing up.
Searching youtube downloader, many people would be fine with some of the ad covered but totally functional sites that pop up on Google. I use some of them every day for quick conversion tasks. I don't want any youtube-dl result. The average users don't either.
Download firefox? What's that? All the top links are fine? No one's looking at the 7th listing for a simple query to download a program.
Why do wider tires have better grip? .. what, sites like roadandtrack, prioritytire, reddit, some physics and stackexchange sites aren't good enough? they are.
The Vancouver snow report one also. Lots of major news sites. Some weathernetwork and almanacs. All totally acceptable results for a sort of variable question.
blah blah this is just a hate on for Google and a HN/nerd view of the world that the average user is nowhere near living in.
They are if the first six results are SEO bullshit. Which is the de-facto state of affairs for Google today: advertising traipsing around as search.
Whole article is rambling and silly and assuming.
Since ddg uses bing, does anyone know what is happening here at bing? It looks like google results are similar.
- "truthsocial trump" works
- "trump truthsocial" doesn't work
There's a power tools review/news site that returns zero hits for the actual domain when searching its name (which is the same as its .com address). While for some domains even searching using the `site:` parameter will give far fewer results when paired with a query than just searching the domain name + query sans the TLD (the router firmware site openwrt.org is among such).
It's a mess and reporting it hasn't any difference ime in the past 3 years. So I'd be reluctant to say irrelevant results are due to censorship unless there was more evidence.
https://www.amnesty.org/en/latest/news/2022/08/ukraine-ukrai...
I use Google as little as possible because I don't like surveillance advertising but fair is fair.
Anyway, here’s Kagi’s bangs:
You can also make your own bangs.
That said, my point was that the bang directory has a bunch of the most useful sites in each category.
Omitting readable styling doesn't read as "techboi rebellion", it reads as ineptitude and lack of respect for people whose attention you're seeking.
If you actually wanted accurate results you wouldn't use a tool that is literally attempting to read your mind like a fortune teller. It is impossible to know what you want just by the word "snow". Jesus Christ engineers are so dumb.