Rediscovering the Small Web (2020)
neustadt.fr
neustadt.fr
The site and list of blogs is open source, growing steadily by about 10 each day (almost at 15,000 at this point).
Every recent post from sites in Kagi Small Web is indexed and given preference in Kagi Search results.
How it works: https://blog.kagi.com/small-web
edit: The project just had its one thousandth commit!
After opening this web page, I pressed down arrow a few times to scroll the page. At first, I didn't understand why it only scrolled a few pixels.
It looks like there's a scrollable area within a scrollable area. The outermost scrollable area only scrolls a few pixels.
This is a badly designed web page.
I tried to access it. It displays a different web page in a frame, which has an invalid certificate (among other things, it is expired and is for the wrong domain name), and then when I bypass the certificate error, I get a 404 error.
...
Look at whom the crime is a benefit in the end.
In similar spirit, check out https://ooh.directory
e.g.
https://alexsci.com/rss-blogroll-network/
This uses OPML blogrolls to crawl blog-to-blog recommendations. I seeded it with the blogs I follow and various planets (https://indieweb.org/planet) and then recursively followed recommendations to build an organic network. Lots of the content is tech-related, indieweb, and smallweb. It's grown to 17 languages and over 4000 RSS/atom feeds.
As an example, the linked blog has a page here [1] and it was discovered by a recommendation by [2].
[1] https://alexsci.com/rss-blogroll-network/discover/feed-12ac5...
[2] https://alexsci.com/rss-blogroll-network/discover/feed-8ecf9...
Is the aggregate list supposed to update regularly?
To reduce storage size, only the title/link/metadata of latest post from each feed is saved. I run the crawler manually, aiming for weekly, but sometimes less frequently. So this won't catch every post and it lags behind quite a bit. I'm hitting some hosting sites faster than I'd like, especially ones that support custom domain names, so I'm planning on fixing up the rate-limiting strategy before I put it in a daily cron job.
There's a plan for ArchiveTeam to use the RSS feed as another way of discovering blog posts to archive. I don't think it's generally useful to point your feed reader at it as there's quite a diverse collection of content.
Personal blogs with real information just can't be found anymore.
https://search.marginalia.nu/search?query=Scaling+a+Digital+...
I've got better phrase matching in the pipe though, give it a few weeks and it should do this even better.
[1] https://www.reddit.com/r/Blogging/comments/i8fmuc/%CA%95%E1%...
Might I suggest (in the interest of privacy) that you give donators the option to use a Silent Payment address instead of a naked BYC address? I noticed you have a Monero address as well, so I assume you care about privacy
Google does not easily surface those websites. Social networks suppress posts with links.
The big modern search engines almost have to be intentionally hiding these websites because they're nearly impossible to find without using an alternative engine like wiby.me or search.marginalia.nu.
Exactly that "surfing" or "webring" or "stumbleupon" style of actually browsing in a larger content rather than searching or push-promote within that pile of content.
If I need to find out what vodka to buy I Google with site:reddit.com and pick the post that's obviously written by alcoholics. The small web can't touch that.
Like my blog has literally 0 SEO and you’ll never find it, but a friend of mine has a blog where he does not post very often, but spends a lot of effort on SEO and it’s very easy to run into his blog.
The SEO meta destroyed small blogs.
To fully understand Google you need to look at them not as a service that brings websites to people, but directs people to websites.
-- Sergey Brin and Larry Page
http://www.zdnet.com/article/google-advertising-and-search-e...
So what we observe in the deterioration of Google search was predicted by its creators, who made the deliberate decision to let this happen by accepting advertising.
Google doesn’t need to downrank small sites, it just happens.
Maybe it’s just semantics
Though as noted, this is not in Google's economic interest to do.
You can drag and drop your entire blog from a single markdown file https://indieweb.social/@xenodium/112265481282475542
You can read the blogs from anywhere, even terminal (no JS needed).
No need to sign up or log in to try it out. I haven't officially launched, but if you'd like to start blogging now, I'll be happy to share an invite code.
Bias disclosure: I have used a text-only client for the last 30 years.
It's basically all the sites and feeds I follow daily with the Hey Homepage built-in RSS reader. You can browse the list and click around, or download it as an OPML file.
RSS = Really Social Sites; OPML = Other People's Meaningful Links
You can see get to some of them here
Collaborative Directory of Geminispace: gemini://cdg.thegonz.net/
But you need a Gemini reader
I discovered it as a young lad lost when playing some RPGs on emulators in the early 2000s
and if you're a front end developer it was apple launching the meta viewport tag in 2007 killed the simple front end.