An article about internet censorship of an archive site that requires an archive site to read!
That’s Brilliant! :D
I am dating myself, but I remember the Internet. It’s rather ironic too that the actual Internet, literal and figuratively free internet, was always doomed from the start, just like AI is effectively doomed in terms of it being some kind of productivity tool or amazing thing you can use to build your own thing, because the logical conclusion is very bad for most people, including the technorati types here that think themselves above it and immune.
If you’re shocked and surprised by the way there Internet was maneuvered into the surveillance prison Flock made itself the poster child of at the hands of Mr Langley, take that trajectory and lay it over AI, but only at exponential rates. You don’t have to believe me, it won’t matter either way, it’s happening, short of the Soviet style implosion of the “American” empire.
Maybe once the planet has been made uninhabitable to biological life there will be an AI robot cybernetic splinter group trying to “reclaim humanity”.
And your AI robot cybernetic splinter group thing sounds a lot like the biblical story of Noah than anything that's based on science or you know facts.
because the empire is too large to control all of its interests.
Could have been avoided - too much money to be made in the various stages of it not being avoided.
Overall I would say civil war is more likely (even given how unlikely it is) than just a robotic uprising that will root out all humanity and make the world unfit for biological beings as the original comment implied.
You can very much host websites on your machine (not the whole internet, but then again you can’t host the most powerful models either) and you can very much download a Wikipedia or Internet Archive.
But most people never will. This idea that everyone will soon be using local models while frolicking happily in the fields is a pipe dream. Besides, local or not, these models can already be harmful in ways the internet took decades to approach.
A local model works even if the Internet is perfectly firewalled or straight-up down. However, you still need electricity and eventually replacement parts, so in an apocalypse may not work either.
A local model works even if the Internet is perfectly firewalled or straight-up down. However, you still need electricity and eventually replacement parts, so in an apocalypse may not work either.
You do understand that was an example, right? Did you know there are other websites you can download? Anna has one that is pretty big and full of information. I’m pretty sure the AI labs know about it too.
> the Internet Archive is way too large to download on a consumer machine.
The Internet Archive isn’t a single zip download. You understand you can pick and choose what you want to download, right? And also, that it was just another example of many?
You couldn’t snapshot all, but you could definitely get a taste :)
Now it’s in the cloud on a damn webpage.
With recent RAM and storage price disruptions, I feel like we’ll go to go back to an equivalent of everyone having a dumb Wyze terminal with the real “computer” in the cloud… Only difference is the dumb terminal will have color, be wireless, and portable.
And I really hope I’m wrong…
It does suck. Though Android’s owner does appear to be afraid of Graphene OS so maybe there’s hope.
GTK, QT, FLTK or TK apps coexist peacefully on linux the same way apps made with Cocoa, JUCE or whatever own internal proprietary toolkit apps run and coexist on MacOS.
"For" the military is close enough, although "defense research" would probably be more accurate than "military" given that initial users were centered around large universities.
From the start the DoD had issues with the free and open culture from the userbase that skewed towards academic types, later splitting off into a restricted MILNET after a few years.
The reason I keep repeating this whenever I see this is because I've seen idiotic mislabeling and mischaracterization materialize as idiotic legislature or idiotic activism based on the former. Someone invented the term "AI datacenter", and now there's a swiping movement of idiots fighting... datacenters. Same as with 5G before it.
Please. Please. Don't blame the Internet for something it has absolutely nothing to do with. Your misattributed criticism is going to be amplified by thousands of automatic and semi-automatic morons and, who knows, maybe people will start cutting cables or attack ISP crew as a result.
The platform needs to be a monopoly or an oligopoly. A highly distributed platform would be a lot more resilient to the problems created by modern social media because it would be too hard for a single organization to manipulate the news reported by thousands of independent publishers.
The platform needs to financially sustain itself in such a way that the users don't directly contribute to it, otherwise it would be way too easy for users to "vote with their wallet".
The platform use should be allowed (or even encouraged) by state or other important service providers, thus making its use mandatory for the holdouts / those who would normally not use it on conscientious grounds.
Perhaps there are other aspects that I've overlooked.
Phone apps are, in general, an even worse alternative than the original Web because of being either completely closed-source, or even if not, not really allowing simple viewing and modification of the program behavior by the users. The way they are created and distributed amplifies the monopoly aspect and the government / service provider sanctioning aspect.
But, I think, that the problems we see now in Web can be partially addressed technologically and partially legislatively. It doesn't have to be this bad.
On technological level, if social media was available as an Internet application, i.e. as an additional service from your ISP, which would remove the need to sustain it with ads, it would make it less interesting for companies like those that run modern social media platforms to manipulate it.
On legislative level, there should be some form of anti information monopoly laws that would make it illegal to have such an outsized effect for social media platforms and would allow the court to require that a platform is split into multiple platforms. The laws that assign responsibility for the published content and its distribution should also address the problem of mis- and disinformation proliferation through such platforms. Even up to the point that they make it impossible for such platforms to exist (today it's common to blame the lack of moderation on the poor quality of content).
That's the name of the site covering the block.
AI helped me and thousandas of other people already to actually build their thing! :-)
Here[0] is an archive.org page which archives an archive.is page which archives the original article about internet censorship of an archive site that might require an archive of an archive site to read (given that people of spain cant now view archive.is in the first place or might have some difficulties doing so)
[0]: https://web.archive.org/web/20260920074416/https://serjaimel...
(If someone is perhaps interested and wants to see me talk more about what this is, then please read the blogpost that I had made for more info: https://smileplease.mataroa.blog/blog/htmlpipe-and-how-we-ca...)
It seems that all of internet archive was down and I don't think that it was because of me (I hope), maybe just some other issues and some sort of just a chance that internet archive got offline at the exact same time that I had shared this thing here.
but when I had opened up hackernews and I read your comment, I genuinely felt myself also as what did I DO.
It is back again now for what its worth, whew what a relief but folks please remember to donate to internet archive for its posterity :-)
Can you please tell me more details so that I can hopefully try to fix it in the future. Did you try to open up the archive.org link of the website or the website itself (which is just a gh page: https://serjaimelannister.github.io/htmlpipe/?https://ppng.i...)
I would like to know more so that I can hopefully fix that for the future, have a nice day :-D
Anyone know of any good reliable substitutes?
For me, archive.today, archive.is, archive.md, archive.ph, etc. are _not reliable_ for a number of reasons
But some archive.today users who comment on HN cannot seem to accept that archive.today may not work for everybody else
NB. Archive.today is not a "substitute for archive.org". Archive.today does not do www crawls
As for archive.org, I know of a number of alternatives but each is generally less reliable and/or less comprehensive than archive.org
Comman Crawl, i.e., downloads from data.commoncrawl.org, is reasonably reliable but not as comprehensive as archive.org. CC is not a reasonable substitute for archive.org's CDX service. The CC CDX endpoint, index.commoncrawl.org, historically has been easily overwhelmed and unreliable
As for archive.today alternatives (no crawls, only user-submitted URLs), ghostarchive.org seems well-designed but not used much. No CAPTCHA, HTTPS and Javascript are optional and HAR files are provided. Whether it gets blocked like archive.today sites I do not know
NB. Archive.today users may be using archive.today not as an archive but as a lazy man's solution for "paywalls" (Javascript annoyances)
Where that's the case, comparsions to archive.org or other archives that are derived from crawls are inappropriate
Use our Parquet index.
It's also worth noting that archive.org downloads all of our crawl data and adds it to the IA Wayback Machine.
https://superuser.com/questions/1346634/modern-browser-with-...
Although I do like the comment at the bottom that states the problem as disabling SNI not encrypting it. Encrypting SNI/ClientHello is over complicated, which is why ESNI was flawed and (allegedly) why Cloudflare disabled it. The solution to the plaintext SNI problem is to not send SNI (I don't send it unless necessary)
There is an alternative non-TLS method of encrypting traffic, per packet, that allows hosting multiple websites on the same IP. I use it in the homelab. It proves that TLS and SNI is not the only way
There's also a popular archive of www content that does not require SNI. It's older and larger than Cloudflare
The problem with software like Firefox is that it automatically sends SNI to every website no matter if SNI is required or not. The superuser thread mentions a Firefox add-on that no longer works. If Firefox is open source then why not just edit the code and recompile
Clearly, Mozilla is not going to provide a solution. It would rather add support for ESNI and then ECH as opposed to giving users an option to diable sending SNI
Mozilla is pro-surveillance advertising, Cloudflare is pro-surveillance advertising
Fortunately, not every HTTPS website is hosted on a shared IP, not every HTTPS website requires SNI. And popular web browsers derived from Mozilla and Google are not the only user agents
A couple of ways to not send SNI
1. Use an SSL client, e.g., openssl s_client, bssl client, etc.
2. Use a local forward proxy, e.g., stunnel, haproxy, etc.
Even if Cloudflare enables ECH across all the websites it controls, and we have been waiting for years, there is still the issue of SSL termination by Cloudflare. For many of those sites, all the TLS traffic, not just the SNI, is available as plaintext to Cloudflare and to whomever Cloudflare, a US corporation, may or must share it with
The strange thing is I use the same DNS on both (which exits on my fiber connection, even when I'm on mobile), so they must be blocking it another way.
Fortunately, I already have a VPN anyway :)
Fucker alters archive.today snapshots, uses the archive.today to do DDoS attacks.
https://arstechnica.com/tech-policy/2026/02/wikipedia-bans-a...
Overall, the hate campaign against archive.today is based on them messing with a doxxer, and I cannot morally support doxxing of anyone. So please, serve me the evil ddos version of archive.today when I visit, and I'll turn a blind eye. :3
Another reason could for example be tied to their identity, which we do not yet know, as they have not [yet] been doxxed.
(Also, idk what pronouns they use, some people in the thread seem to refer to them by "he", is that actually something they refer to themselves as on the record, or rather a typo or 'borken Engrish'?)
Vs
One has paywall, one doesn't...
Sorry man, people will keep DDoSing some random guy's blog. It's a service problem lol
But despite this I feel like what he did was very counterproductive to this goal. If he did nothing, people would just forget about it. Wikipedia would still use his service because they don't want to deal with copyright.
Besides if western intelligence really wants to find this guy, they could probably do it without the journalist's help.
Like if a very infamous site, KiwiFarms can have their own archiving engine, why cant multimillion dollar site like Wikipedia can't? He's absolutely correct that Wikipedia is a freeloader that doesnt want to deal with copyright.
However, I'm not sure if the archive guy is paying it out of his own pocket. It's a massive ad-free site servicing tons of requests everyday, and it has a very impressive technical abilities.
That has to be quite expensive, like he must've some seriously rich sponsors. And his sponsors might be okay with Wikipedia using the site, but this is all just my guess lol.
In the end, it's also perfectly valid for him to refuse servicing Wikipedia: it's his site.
It was completely counter-productive too, several orders of magnitude more people learned about the blog post via the DDOS controversy than ever would have without it. And plants a seed of doubt that anything they archive could have been silently modified to satisfy a personal grievance.
Edit: eeh, that's a different site than the one mentioned by the submission, why it wouldn't work?
Edit2: Also, the error is wrong and server-side. If your ISP is trying to block the website, you'll get either a tls error that the certificate isn't matching, or for non-https sites you get redirected to the scary "You're contributing to something illegal blah blah".
Here you go, this should work: https://web.archive.org/web/20260920074416/https://serjaimel...
(an archive.org page which archives an archive.is page, might take some time to load though)
(this uses htmlpipe, something I have made :-D)