I pay Kagi, they give me search results. It’s not an ecosystem. None of their competitors are all that much of an ecosystem, either.
If Kagi goes out of business tomorrow it’s really not my problem, with the exception of the prorated balance of the subscription.
The only thing that really matters is that running searches in Kagi is a better experience in Kagi than Google or Bing or DuckDuckGo or most of the rest. That’s what makes me pay for it.
I don’t really need to know how the sausage is made (except for the privacy aspect).
The speed with which many of us respond when a tech company does something really stupid is a key reason that things are not even shitter.
If Microsoft hasn’t managed to build a meaningful search index after 17 years of sinking billions in it, there’s a good chance that no one ever will.
[0]: https://blog.kagi.com/waiting-dawn-search (cf. section _The problem: A search monopoly_)
I'll continue to support them because it's a great experience.
https://kagifeedback.org/d/4727-option-to-choose-or-exclude-...
I ended up quitting Kagi over that and built my own search engine for my own use. But if the comments here are correct that Kagi are trying to build their own index that replaces all the others, that would be great news. It seems lots of people still like Kagi regardless of that too.
For anyone trying to build their own, turns out SQLite can take you much further than it seems you should ever be able to. A metasearch layer fills in for everything else.
I struggle with indexing too. I don't do any crawling, just indexing from sitemap.xml files or URLs I add to the queue manually. Also I'm indexing from my own residential PC, not a hosted VPS. Something that helped me was Cloudflare's new crawler API. For all the sites where Cloudflare blocks you, just use their API instead:
https://developers.cloudflare.com/browser-run/quick-actions/...
https://developers.cloudflare.com/browser-run/quick-actions/...
Also for a quick index bootstrap, the Curlie database can be useful. It's the old DMOZ directory, the open Yahoo competitor, with 1.4 Million websites. A lot of the links are now very outdated, but it can be useful, and it's less than 500MB when converted to an SQLite database.
More details are at: https://help.kagi.com/kagi/search-details/search-sources.htm...
Also, for every query, it shows a short infoline telling the index distribution.
https://web.archive.org/web/20240901052114/https://help.kagi...
I'm aware of their Teclis index, I used to be a paying API customer. Teclis is very small. It's primarily an index of indie blog websites (smallweb) mostly crawled via RSS feeds. That itself is a very cool idea, and a great supplemental index to have! But it isn't the kind of index that could ever replace their Brave / Yandex / SerpAPI dependencies. The Teclis API is even supplemented with results from Marginalia Search, a much larger index created by one person with less funding.
There's some technical info here on how Teclis is indexed, eg Readability.js for content extraction & Elasticsearch for the full-text indexing:
Kagi's short infoline about index distribution sounds new to me though. I would love to see a blog post from Kagi about that & where they're at with query coverage.
I might try it again sometime.
I’m just one person, and that’s just one incident, but it happened pretty early on after I switched to Kagi and had a pretty big impact on my perception of it.
I had been trying to get away from Google for many years, and with other options like DDG, I always felt I needed to go to Google for various searches, but so far I haven’t been back to Google at all in several years now.