Twitter now requires an account to view tweets
techcrunch.com
techcrunch.com
"This will be unlocked shortly. Per my earlier post, drastic & immediate action was necessary due to EXTREME levels of data scraping.
Almost every company doing AI, from startups to some of the biggest corporations on Earth, was scraping vast amounts of data.
It is rather galling to have to bring large numbers of servers online on an emergency basis just to facilitate some AI startup’s outrageous valuation."
> Something went wrong. Try reloading.
on that link. Which, frankly, is hilarious.
Isn't that basically what they did with the API changes?
Which just leads people to scrape.
If they'd made some more reasonable pricing tiers, I would have been happy to pay.
Fetching something as simple as the total follower count from an API shouldn't be more (exorbitantly) more expensive than fetching data from, say, GPT-4. No reasonable person can make an argument for $10c/call pricing.
If it doesn't respect robots.txt, it is unethical.
b) Given Twitter is public, user generated content which they don't own but simply have a license I wouldn't call it unethical in the slightest.
I do a lot of data scraping, so I’m sympathetic to the people who want to do it, but violating the robots.txt (or other published policies) is absolutely unethical, regardless of the license of the content the service is hosting. Another way of describing an unauthorised usecase taking a service offline is a denial of service attack, which (again, if Musk’s description of the problem is accurate) seems to be the issue Twitter was facing, with a choice between restricting services or scaling forever to meet the scrapers requirements.
Personally I would have probably tried to start with a captcha, but all this dogpiling just looks like low effort Musk hate. The prevailing sentiment on HN has become so passionately anti-Musk that it’s hard to view any criticism of him or Twitter here with any credibility.
The only reason these websites and platforms aggregate any content at all is because they're effectively giant public squares.
Musk is trying to have his cake and eat it...
(Clearly it's not a public square, but his position is incoherent).
It is NOT legal to install cameras that record everyone's conversations, much less sell the laundered results.
Pre-2023 people went on Twitter with the expectation that their output would be read by humans.
A traditional search engine is different: It redirects to the original. A bastardized search engine that shows snippets is more questionable, but still miles away from the AI steal.
Expectations =/= reality. And the reality is that bits have been reading comments for over a decade.
When you have financial incentive to build your business on someone’s data and you scrap literally millions if not billions of pages - it’s unethical.
This data is often of great public value. I track conversations around a social issue as part of my work for a non-profit.
I'd counter it's unethical to prevent people from accessing this data.
> great public value
Having been to twitter mostly through the most recent prominent war, man the signal to noise ratio is really low even when being careful about who to follow and who to block. There is so much disinformation, bad takes, uninformed opinions presented as facts, pure evil, etc.
So I guess it could be used for training very specific things or cataloging the underbelly of humanity but for general human knowledge it’s a frigging cesspool.
I do not use the Twitters myself, and actively discourage others from doing so. Sends people bonkers.
We were actually gearing up to switch to paid accounts as we found use cases that could subsidize these efforts... And then the starting price for reasonably small volumes shot up to like $500k/yr.
We have robots.txt. If Google doesn't respect that, it's unethical. Don't you think so?
Not that the people you want to respect that would
Especially since they're not moderating things or anything.
Agreed. However, it's probably covered by their terms of service.
Same thing with the recent reddit kerfuffle. I'd have much preferred a Usenet 2.0 instead of centralizing global communications in the hands of a handful of private companies with associated user-hostile incentive structures.
I don't despise Musk at all. Don't agree with him on everything, but he is a genuine and interesting person.
Then they followed it up with the old hacker news chestnut of 'whatabout the west'
I don’t think it’s a particularly nuanced point to you
What point and why do you keep saying 'nuance' over and over while giving zero actual information? What are you trying to say and what evidence is there?
Let's think about this super hard. What is the justification for an unprovoked genocidal war? Why are you defending putin?
however ignorant it might be.
Show me where you get your information, lets see the source of this nonsense.
2. If you actually want to hear my opinion, then the realm of geopolitics + good old-fashioned hate of the US government does a number on people's logic, so we get what I can only charitably describe as a parade of non-sequiturs, whataboutisms and other fallacies. And so it can be useful to frame it in the simpler terms, for example you could hardly find anyone even on this site who would condone the forced takeover of parts of people's homes. Literally the same is happening on a scale of the countries.
It cuts right through the bullshit and frequently exposes a lot of hypocrisy, hate, self loathing, racism, etc.
That's an extremely dumb take
Your nonsense straight out of the a Soviet propaganda book doesn't work on me.
Go ask people from Eastern Europe how they felt about 45/50+ years of Soviet imposed governments and regimes.
The West has done awful things but in what way do they excuse the attempted ethnic cleansing of Ukraine?
You’re probably smart enough to understand that out of spite and regret of your country’s history with the Russians your countrymen have more motivation than many others to judge the Russian efforts without any further investigation into the matter.
The same applies for myself, since I’m Finnish. It’s almost sad to see how people abandon all reason and critical thinking skills because of some ingrained belief that “Russia bad”. All of my knowledge of the human nature leads me to believe that they’re no more bad than the next people, and that they probably have some motives to go to a taxing war that we don’t really understand here in the west - seeing as the first casualty in war is the truth.
It's interesting how you assume that most people despise Musk.
But why not just... provide it? Charge however much for a box of hard drives containing every publicly-available tweet, mailed to the address of buyer's choosing. Then the startups get their stupid tweets and you don't have any load problems on your servers.
Wow, never thought of it that way before. Kinda hit me hard for some reason.
I think "AI is takin' ooor contents!" is a convenient excuse to tighten the screws further. Having a Boogeyman in the form of technology that's already under worried discussion by press and politicians is a great way to convince users how super-super-serious the problem must be, and to blow a dog whistle at other companies to indicate they should so the same.
It's no coincidence that the first two companies to do this so actively and recently are both overvalued, not profitable, and don't actually directly produce any of the content on their platforms.
In other words, the root problem is incompetent management, not any technical issue.
Don't worry though, the legal system is still coming for Musk, and he will be forced to cough up the additional billions (?) he has unlawfully cheated out of a wide assortment of counterparties in violation of his various contracts. And as employee attrition continues, whatever technical problems Twitter has today will only get worse, with or without "scraping".
(And IMO, that perceived legitimacy was unfounded for both HTTPS and the blue check before both were easy to get, it's just that the bar had to drop to the floor for most people to realize how little it meant.)
He got laughed out of a Twitter call thing with lead engineers in the industry for saying he wanted to “rewrite the entire stack” and not having a definition for what he meant.
Doomed or not, Musk is terrible at just about everything he does and Twitter is no exception
The key benefit to a cache is that a small set of content accounts for a large set of traffic. This can be staggeringly effective with even a very limited amount of caching.
Your options are:
1. Maintain the same cache size. This means your origin servers get far more requests, and that you perform far more cache evictions. Both run "hotter" and are less efficient.
2. Increase the cache size. Problem here is that you're moving a lot of low-yield data to the cache. On average it's ... only requested once, so you're paying for far more storage, you're not reducing traffic by much (everything still has to be served from origin), and your costs just went up a lot.
3. Throttle traffic. The sensible place to do this IMO would be for traffic from the caching layer to the origin servers, and preferably for requesting clients which are making an abnormally large set of non-cached object requests. Serve the legitimate traffic reasonably quickly, but trickle out cold results to high-demand clients slowly. I don't know to what extent caching systems already incorporate this, though I suspect at least some of this is implemented.
4. Provide an alternate archival interface. This is its own separately maintained and networked store, might have regulated or metered access (perhaps through an API), might also serve out specific content on a schedule (e.g., X blocks or Y timespan of data are available at specific times, perhaps over multipath protocols), to help manage caching. Alternatively, partner with a specific datacentre provider to serve the data within given facilities, reducing backbone-transit costs and limitations.
5. Drop-ship data on request. The "stationwagon full of data tapes" solution.
6. Provide access to representative samples of data. LLM AI apparently likes to eat everything it can get its hands on, but for many purposes, selectively-sampled data may be sufficient for statistical analysis, trendspotting, and even much security analysis. Random sampling is, through another lens, an unbiased method for discarding data to avoid information overload.
https://www.wired.com/story/twitter-data-api-prices-out-near...
Researchers are complaining that it's far too high for academic grants. Probably true, but that's no different from other obscenely priced subscriptions like access to satellite imagery (can easily be $1k for a single image which you have no right to distribute). I'm less convinced that it's impossible for them to do research with 50 million tweets a month, or with what data there is available. Most researchers can't afford any of the AI SAAS company subscriptions anyway. Data labelling platforms - without the workers - can cost 10-20k a year. I spoke to one company that wouldn't get out of bed for a contract less than 100k. Most offer a free tier a la Matlab in the hope that students will spin out companies and then sign up. I don't have an opinion on what archival tweets should cost, but I do think it's an opportunity to explore more efficient analyses.
That said, there are concerns with data aggregation, as patterns and trends become visible which aren't clear in small-sample or live-stream (that is, available in near-time to its creation) data. And the creators of corpora such as Twitter, Facebook YouTube, TikTok, etc., might well have reason to be concerned.
This isn't idle or uninformed. I've done data analysis in the past on what were for the time considered to be large datasets. I've been analyzing HN front-page activity for the past month or so, which is interesting. I've found it somewhat concerning when looking at individual user data, though, here being the submitter of front-page items. It's possible to look at patterns over time (who does and does not make submissions on specific days of the week?) or across sites (what accounts heavily contribute to specific website submissions?). In the latter case, I'd been told by someone (in the context of discussing my project) of an alt identity they have on HN, and could see that the alternate was also strongly represented among submitters of a specific site.
Yes, the information is public. Yes, anyone with a couple of days to burn downloading the front-page archive could do similar analysis. And yes, there's far more intrusive data analytics being done as we speak at vastly greater scale, precision, and insights. That doesn't make me any more comfortable taking a deep dive into that space.
It's one thing to be in public amongst throngs or a crowd, with incidental encounters leaving little trace. It's another to be followed, tracked, and recorded in minute detail, and more, for that to occur for large populations. Not a hypothetical, mind, but present-day reality.
The fact that incidental conversations and sharings of experiences are now centralised, recorded, analyzed, identified, and shared amongst myriad groups with a wide range of interests is a growing concern. The notion of "publishing" used to involve a very deliberate process of crafting and memoising a message, then distributing it through specific channels. Today, we publish our lives through incidental data smog, utterly without our awareness or involvement for the most part. And often in jurisdictions and societies with few or no protections, or regard for human and civil rights, let alone a strong personal privacy tradition.
As I've said many times in many variants of this discussion, scale matters, and present scale is utterly unprecedented.
One of the things you could do is reduce the granularity. So instead of showing that someone posted at 1:23:45 PM on Saturday, July 1, 2023, you show that they posted the week of June 25, 2023. Then you're not going to be doing much time of day or day of week analysis because you don't have that anymore.
Though I've thought for quite some time that making the trade and transaction of such data illegal might help a lot.
Otherwise ... what I see many people falling into the trap of is thinking of their discussions amongst friends online as equivalent, say, to a discussion in a public space such as a park or cafe --- possibly overheard by bystanders, but not broadcast to the world.
In fact there is both a recording and distribution modality attached to online discussions that's utterly different to such spoken conversations, and those also give rise to the capability to aggregate and correlate information from many sources.
Socially, legally, psychologically, legislatively, and even technically, we're ill-equipped to deal with this.
Fuzzing and randomising data can help, but has been shown to be stubbornly prone to de-fuzzing and de-randomising, especially where it can be correlated to other signals, either unfuzzed or differently-fuzzed.
> I’m still confused as to how a non-profit to which I donated ~$100M somehow became a $30B market cap for-profit. If this is legal, why doesn’t everyone do it?
https://twitter.com/elonmusk/status/1636047019893481474
> I donated the first $100M to OpenAI when it was a non-profit, but have no ownership or control
Because it was thereby starting to compete with his own for-profit.
He was even asking for a moratorium for 6 months so his company could catch up.
He tried to pressure them to make him a CEO, they refused, so he said "no money then, go bankrupt" and quit the board. They made a deal with Microsoft and survived.
Now he's pissed.
Then in 2022, they blew up and Elon's been spitting venom at them ever since as he missed his chance.
It sounds like Elon doesn't "get" the open web
The licenses, compensation models, law, technical solutions, attribution, security and privacy all need time to catch up. Regulation has a role to play as its a bit of a free for all right now.
The irony of Elon mentioning “outrageous valuations” though!
I'm strictly anti account, so he just lost me as audience. The next walled garden after Facebook and Instagram that won't ever see me again.
However if you mean implementing some even worse obfuscation (kind of like FB putting parts of words in different divs etc) that is not really compatible with the situation that this needed to be done as more of a temporary emergency measure. And PoW doesn't sound reasonable because it sets mobile devices against the scraper's servers. If all of this was just so easy, scraping would be dead. Good that it isn't.
Scraper servers and mobile devices have different access patterns though. I I'm reading tweets then I'm fine waiting 1 second for a tweet to load. Page load times for this kind of bloated stuff are super slow anyway, meanwhile my mobile could spend a second or two on some PoW. But if you want to large-scale scrape, you suddenly have to pay for 1bn CPU seconds. And this PoW could even keep continuously increasing per IP. 0.1% with every tweet. Not noticeablr for the casual surfer sitting on the toilet, neck-breaking for scrapers.
> If all of this was just so easy, scraping would be dead. Good that it isn't.
Small-scale scraping could still be provided through API access or just a login.
The reason they are not doing the "easy" thing is that they don't see a need (yet, perhaps). Just get an account, they'd say, and they are right. It works for Instagram too, except for some weirdos who nobody really cares about.
>The reason they are not doing the "easy" thing is that they don't see a need (yet, perhaps).
This argument just doesn't make any sense. Twitter notes that this is hurting them. Previews in chat apps, just clicking links in non-loggedin contexts is are broken. I feel like you just predict that this will turn out to be more accepted in the near future and become more a more permanent decision, which you don't like.
Most baffling is mobile reddit, where it takes like 6 seconds to load. Do they want us to use their crappy app, or they just dont care?
For non-auth use, I rather wait for 1 second than not have any access at all. Which is the current state of affairs.
Also a side note: distributed Web crawler is not unheard of these days, as well as residential IP proxies. Meaning the effectiveness of Proof of Work model maybe also limited.
These systems tend to treat residential proxies as normal users, and puts less restrictions on them. On the other hand, if the IP address belongs to some (untrusted) IDCs, then the system will enable more annoying restrictions (say rate limits etc) against it, making scraping less efficient.
Correct me if I'm wrong, but surely throttling scrapers (at least ones that are not nefarious in their habits) is a problem that can be mitigated server-side, so I find it somewhat galling that its the excuse.
No matter what you do, this will cost server infra. That's Musk's argument for disabling access altogether.
Therefore it would make sense to have a solution which burdens the client disproportionately in relation to the server. A burden so low for the casual user that it's negligible but in aggregate, at scale, would break things. Which is what he wants.
Looks to me like both reddit and twitter are using the wedge to rather increase the height of the wall of their gardens and kill 3rd party development as opposed to genuinely trying to license bulk-users appropriately.
You're gonna need to license api keys so you're already identifying consumers and there's your infra which you need anyway. At which point you can throttle anyone obviously abusing whatever free/open-source tier offering you give out as standard.
This works far better when the items requested are small in number but large in volume (that is: a large number of requests against a small set of origin resources). When dealing with widespread and deep scraping, other strategies might be necessary, but these aren't impossible to envision.
Specifically permitted scraping interfaces or APIs for large-volume data access would be another option.
Of course, there's the associated issue that data aggregation itself conveys insights and power, and there might be concerns amongst those who think they're providing incidental and low-volume access to records discovering that there's a wholesale trade occurring in the background (whether that's remunerated or free of charge).
Many people who aren’t that privacy conscious would however object to lots of companies, big and small, sucking their content into their databases for their own uses, then republishing after it’s passed through a few AI models.
I think this points out that AIs clearly do not work like human brains. Human brains do not need all of the content of humanity to produce a replica of art station mediocre.
Do they have a choice? A handful of corporations have captured all the network effects. If you need to reach an audience to do your job or find your "friends", what other choice do you have but to give your data to them?
If they aren't close why do you care?
Alternatively, what would make more sense is to participate in communities gated by these networks but then it's your choice to be there.
I don't personally have this problem, but my observation is that most social relationships are somewhere in between closest friends and don't care.
My own concern is more about participating in professional, neighbourhood, civil society or political communities. Choosing not to be where they have decided to congregate means not being able to do my job and not making my voice heard where many decisions affecting me are taken.
Actually I’m surprised this took so long to do, and in the light of doing so shows that perhaps Twitter was sold for its existing content rather than existing or active user base.
Not a very significant distinction: if active users stopped posting, scrappers wouldn't have much of a reason to keep scrapping.
Age may be included as part of the training but they generally want to suck up as much data as possible.
It's shared ownership. You own it, but give Twitter non-exclusive permission to also use it.
This is why news agencies request permission from a twitter account before sharing a picture they took.
I can own a picture, but I can't place it on the NY Times website.
It may seem weird weird to compare useful content to a used napkin, but hey, successful business founder stereotypes do quite often involve have an idea written on a napkin…
Yes, it’s a mistake to rely on social media content remaining up forever, agreed. That’s separate from ownership. Backups are important even for data on a hard drive you physically own, since hard drives can fail or be damaged or lost.
I'd expect those that require real time data, such as stock market bots or sentiment data providers, to scrape twitter (if they don't provide the data by other means, for example the "firehose", which is another great way to earn money).
None of this makes much sense.
Also, it's much more complicated than it seems. The web works because the data is public. You cannot think of it as "my data". (Especially not twitter, since it is really their users'!) Twitter is not higher quality data than any other web page.
If we accept that thinking, every home page would require a login to see that specific company's phone number of opening hours. Those pieces of data are also valuable, in the right circumstances! And then the web would either not exist or the account system required would be so wide spread that accounts would carry no value and the system would become useless.
What do you think the webs value was imagined to be in the “80s”?
There was no web in the 80s so not sure what values you refer to, or how they are relevant to today's businesses.
Of course, none of those people own Twitter, and it may well be more valuable to its owners if it does require a login.
I can think of a few possible reasons. They might want more up-to-date info, or they might have no real developers and the scraper was created by a business guru who prompted ChatGPT and didn't understand the code that came out.
Given what else Musk has asserted about Twitter, and how often former or current Twitter devs have contradicted him, it may not even be what Musk said.
> Twitter is not higher quality data than any other web page
Eh, depends how much you can infer from retweet, favourites, etc.
Won't be the only such site, but it's probably better training data than blog posts are these days.
But yeah, I absolutely agree that Twitter doing this caused a lot of damage to any orgs, corporate or government, which wanted to be public, anything from restaurants announcing special offers to governments issuing hurricane warnings. Twitter isn't big enough to assume everyone has an account, like Facebook is.
This is just kicking the can down the road.
We have a major structural problem now. We want data to be free and machine readable, but no startup (and even a giant like Twitter) can afford the server cost to withstand all those machines.
But then so is Twitter. They don’t produce any content whatsoever. The data they are having a fit about is not theirs, it’s been volunteered by the users. It’s the same line Reddit is pushing, and it’s bullshit. AI companies scraping the web is no more unethical than Google doing it.
Taking/scrapping/stealing that data out of said platform for the benefit of your over-hyped “disruptive” startup - and implying that others should give you all for free - is the issue.
The Twitters and Reddits need to be careful here when complaining because without users generating free content, they also have no business.
I recall that email at one point was 90% spam.
Also, Twitter is a public platform. Twitter didn't generate comments, and people posting on a public account are indirectly subject to public viewing. Not much different from being indirectly recorded in a public park
Including you and me. WE built part of the infrastructure.
There's one major player making money off of other peoples content here, and that's Twitter. Why are they ok doing that, but not anyone else?
Scraping data from a public website is not "stealing". It might be a violation of the terms of service, but then you have the whole issue of click-through (formerly shrink-wrap) licenses and contracts of adhesion.
If someone isn't vetting you and potentially signing you to a more meaningful contract before giving you access, for free, to data, then using that data for any purpose whatsoever (except republishing it or derived works, which might, depending on the nature of the data or the derived works, be a violation of someone's copyright) is so far from "stealing" that using that word is wrong, and I suspect intentionally inflammatory.
That's why Elon limited access, rather than going to the police to file charges for theft, or suing over copyright violation or breach of contract. Not to say he absolutely couldn't do the latter, but it's hardly a clear win.
1. Users or one might say content creators don't own their data. Not just do those platform owners make a lot of money with the content (which they have a license to, as per site ToS) but then you have third parties scraping it now for commercial products. Using the data to train models that are then sold back to some of the same social media users who produced the content for free in the first place wasn't a thing until very recently, it used to be a select few doing machine learning research in the past. The laws are lacking behind the tech development and regular internet users are being exploited because of it.
2. It absolutely is stealing in some cases, and even worse. For example when they scrape it for content which they then use to train their bots to impersonate humans. Or on Twitter, there's a very common type of bot that steals content from young attractive female social media users in China, auto-translated to English, to pose as them. If you're in finance and crypto circles they're swarming with these accounts (guess the scammers know their targets).
3. In general this is only going to get worse from here on. LLM are getting better and better. On sites like Twitter you already have no idea if you're interacting with a human or not. But these "AI" can not actually think for themselves, they can only emulate, they can copy other humans. At least so far. So for the sake of making progress and ensuring we can still have intelligent discussions and find novel ideas online, it's imperative to have a way to keep the machines out. Social media must become sybil resistant or it dies in a vicious circle of self-referencing bots ever parroting the same old talking points, or variations thereof. We urgently need human ID!
Nobody is deprived of their property, intellectual or otherwise, when this stuff happens, thus is it not right to call it stealing.
Governments, researchers, and all kinds of third parties have already been scraping every publicly available bit of data possible. There may be an increase now, but it's nothing new. It won't be the end of the society or the end of the internet anymore than AI will.
>have already been scraping every publicly available bit of data possible
Data scraping is limited by economics just like anything else in the world. Storage costs money, someone has to pay for it. Researchers do not have unlimited funds. Some select few governments like the US may have most of the publicly accessible web archived. Keep in mind it's dynamic and requires massive data infrastructure to pull this off, there's tons of new data coming in daily. Private startups getting in on the action in a big way is a relatively new phenomenon, this used to be limited to enterprises with a specific purpose. Now everyone and their 4chan cousin are experimenting with their own deep learning models.
Bots aren't people and can't read nor consent. They just consume.
Any page which can be served without first displaying a ToC or other terms which explicitly prohibit access is not protected by a ToC or other license from scraping, as they can be considered a Point of First Contact in each case, as the bot has selected each link from a simple aggregation of all links it encounters (each interaction being "new" in essence).
Now it could be argued that ignoring robots.txt is an explicit contravention of norms and standards which could be viewed as a violation of an implicit licence, but there is no law requiring adherence to robots.txt and thus no mandate that a program even look for it iiuc.
Even more applicable, this is like saying that a person walking down the street can't have a camera and take a picture of the front of the building....
Because the page you land on when entering a url is in fact little different than a store front, with the associated signage and access points defining how a person or automated device may interact with that business.
If you want to have it different then you have to actually put everything behind a locked door with no window, right?
Reddit is a cesspool of misinformation and phallic references.
I guess there may be certain subreddits/tweeters that someone might want to train on, but I dont understand why.
I wish it would stays that way though so that I wouldn't land on the site by accident.
[1] https://www.datacenterdynamics.com/en/news/twitter-pays-goog...
Congrats you played yourself.
Twitter is maybe at the dog stage. Perhaps it’ll die.
This smells of another failed Musk experiment at twiddling with the knobs to increase engagement, to me.
The legal ambiguity comes from the question of whether LLM outputs are a derivative work of the training data. I expect that they aren't, but anything can happen.
If these data scraping operations are as sophisticated and determined as he claims this measure is insufficient and actually it really hurts Twitter far more than it helps. Case in point: we stopped sharing Twitter links because when you click them in most iOS apps it opens up an unauthenticated web view and presents you with a login screen. So we just collectively decided “ah ok no sharing Twitter” and moved on.
I’m sure there are companies scraping Twitter. I just don’t buy that it’s as big of an issue as he claims it is, and that preventing people from viewing tweets without logging in is a way to mitigate against that (I’d first look at banning problematic IP addresses first, personally).
To me it’s either:
1) a very poor and very temporary mitigation against scraping, that could be bypassed with a bit of effort
2) an experiment in optimising metrics - Musk sees lots of unauthenticated users consuming Twitter, tries to steer them into signing up
3) it’s all just a big mistake
Option #2 makes the most sense to me, but frankly none of them are good
It can add nontrivial load.
Suppose 1 million people are accessing Twitter at any given time. An actual person might only be making 1 request / second. That's 1 million requests / second.
Suppose there are 100 AI companies scraping Twitter. A bot like this can make thousands to tens of thousands of requests per second. That's an additional million requests / second.
There are probably more than 100 "AI" companies now, trying to train their own bespoke LLMs. They're popping up like weeds so I can totally see Twitter's load doubling or tripling recently. So sorry, I just don't get the skepticism. Sure it could be a cover for something else, but his actual stated reason seems totally possible.
You don't need to use a bot to do this, Twitter literally did this to themselves through their own buggy code https://sfba.social/@sysop408/110639474671754723
If Silicon Valley was still being produced, this would make for a great episode.
I think that we are already putting too much content into Social Media platforms (HN included). Stuff that we sort of ought to self host because then we would actually own it. But will you even want to run your own sites publicly if they are getting scraped? I guess it’s it really a new issue as such, but I imagine it’ll only get worse as the LLM craze continues to rise.
I was relieved when they started asking for an account this week because now I'll finally be able to break my habit of navigating to Twitter to "see what's happening" only to find a bunch of sports memes, pop music drama, or right-wing trolls pretending like Hunter Biden is the most nefarious person on the planet.
Imo, Elon is lying and he locked everything down for PR so he would make a headline and frame it like his site's content is _so valuable_ that he just had to take drastic measures to stop AI from training on it.
But hey, I am eternally grateful. As a super-compulsive Twitter user, I read accounts and my own lists after I log out. This fixes it.
Have your "emergency", Elon.
Bad data is devastating in a way a HTTP 401 is not.
A) Twitter would probably move in this direction even if AI companies didn't exist, this is an excuse. Nothing about Musk's Twitter has indicated that he cares about Open data access or anonymous access to the site, and this follows a general trend of closing down the platform to non-monetizable users. Musk has abundantly shown in the past that he would prefer everyone browsing Twitter be logged into an account.
B) "it's temporary" -- how? You don't have a way to stop this other than forcing login. That situation is not going to change next week. To call this "temporary emergency measures" is so funny; there is no engineering solution for this and you're not going to be able to successfully sue companies for scraping Twitter. Put a captcha in front of it? Sure, let me know how that goes.
You going to wait and see if the AI market collapses in the next month?
If this does turn out to be temporary, it'll only be because of migrations off of Twitter and because of user criticism, because Musk is impulsive and bends easily under pressure. But nothing about the situation Musk is complaining about is going to change next week.
Libs like puppeteer is so good these days that it's impossible to tell real users from fake traffic. Most of the blocks are just IP blocks.
Not so, he tasked George Hotz with getting rid of that horrible popup which prevented you from scrolling down much if you weren't logged in, which was added soon before he bought Twitter. When that was removed I rejoiced. But now Twitter's gone 100x in the opposite direction.
To be fair, Musk will regularly pay lip-service to the idea of Open communication. I guess that's not literally nothing, but most large site policies have been in the direction of locking down content.
If there ever was a version of Musk that cared about Open access, it's been a while since that version of him saw the light of day. It's very consistent with his overall behavior to believe that he views Twitter content as being primarily his property rather than a community resource, and that he thinks that scrapers/AI companies/researchers are literally stealing from him if they derive any value at all from data that Twitter hosts.
There absolutely is, if you try instead of whining on internet. People at Vercel have already developed new anti-bot + fingerprinting + rate limiting techniques which look quite promising. I dare say within a year, new tools will be powerful enough to do this easily.
I see where you're coming from, but if Twitter is in a position where it can't roll out those protections right now, given its current head counts, etc... it's not going to be in a position where it can roll out those protections next week. Probably not next month.
So it's less that no one could block companies from scraping Twitter (although anti-scraping mechanisms are probably always going to be a cat-and-mouse game, so I'm not sure that there is ever going to be a perfect easy solution). It's more that if Twitter can't do it right now, nothing is going to magically change any time soon about the situation it has found itself in. And waiting a year (even waiting 6 months) for tools to become available before rolling back this rate limiting would be incredibly self-destructive for Twitter.
The way I see it, they're basically guaranteeing that they will need to roll back these changes before they have a solution to whatever specific problem or irritation Musk is fixated on. They're not going to gain additional engineering capabilities in the next week. And how long does Musk plan to leave rate-limiting in place? A social media site where people can't look at content is just broken.
Fuck.
I guess I'm done with Twitter.
Reddit is in Eternal September. Twitter is login-walled. If HN is next, I'll probably be mostly done with the Internet.
This version of the Internet is starting to suck. :(
It's like twitter and reddit but 15 years ago (and by that I mostly mean it's janky and full of bugs. Just like the web was 15 years ago!)
We are oriented to think that way after ~15 years of the algorithmic engagement-maxi world of Twitter. It always looks like there's a lot happening all the time but look deeper and it's a bunch of people offering their weak takes on hot topics to build their brand.
What was the last thing you remember being must-see sfuff on Twitter?
But I do agree with you, being away from the noise has been really freeing for my mind
The Russia circus last weekend made for some pretty good near real-time intrigue. That being said, I don’t care what crazy thing is happening…I’m not creating an account to hear what’s being said on “The Global Town Square ™”
In theory, Twitter should be that real-time news feed from these kind of events, but it doesn't actually work. Signal-to-noise ratio is just very low.
I never cared about algorithmic engagement.
Before (and during) the engagement Twitter:
- has everyone you needed there, centralised
- has search, where you can find people and topics you care about
Mastodon has none of that: you have to know which server to join and how to find people and topics. Centralization always beats distribution in convenience.
It's kinda like the "I don't care about politics" stand. You might not care, but the institutions you interact with every day certainly do.
This is extremely useful
Is that the best way? No clue. But I'm glad to chip in.
(Note: I'm of the opinion that fediverse-style federation in the context of forums is merely a nice-to-have; the web is already naturally federated, and people should not feel bad if they want to save money/tech complexity/administration complexity by settling for ordinary self-hosted forums.)
Is there a reason for this, though? Need to be able to iterate on features quickly? Maybe not being able to tackle various complexities with the total available resources? Or maybe federation is just inherently expensive?
Why couldn't we have an alternative written in a more performant language/runtime with maybe things like lower quality images/videos or something?
Because performance was not a concern when it was designed - or it could be that it was designed for small communities, and therefore not possible to scale-up cheaply. One of the problem is the caching of pictures from the different instances connected (if I remember correctly) which makes the data storage requirements go up very fast
Instead, it feels like the current Fediverse demands that I make a blind choice to entrust not merely a copy of my content but also my whole future identity to whatever of these current instances looks the most stable/trustworthy at first glance, hoping my choice will be good for 1-5-10-15 years. It's stressful, and then I look into self-hosting, and then I put the whole thing off for another week...
AFAICT I would need to set up a whole federated node of my own in order to get that level of identity-control. Serious question: Is there any technical limitation preventing the admin of an instance from just seizing an particular account and permanently impersonating the original owner?
In contrast, I was hoping/expecting some kind of identity backed by a private asymmetric key. Even if signing every single message would be impractical, one could at least use it to prove "The person bob@banana.instance has the same private key that was used to initialize bob@apple.instance."
"bsky.app" works as a web client for the official "bsky.social" instance, but it also works with the instance I self-host (or any other spec-compliant instance). Likewise, 3rd party clients work with the official instance, and also with 3rd party instances.
However, no key-stealing could possibly happen right now in any case because... the PDS ("instance") holds your signing key - the client never even sees it. Having the server hold your signing keys is very user-friendly, but of course not ideal for security and identity self-sovereignty. In general, the security model involves trusting your PDS (just as you trust your mastodon instance admin, or twitter dot com - the improvements are centered around making it easier to jump ship if you change your mind).
Client-signed posting is something that's not even possible right now, but I believe it's somewhere on the roadmap. If it doesn't happen some time soon I'll be implementing it myself. (I'm writing my own PDS software)
But yes, the protocol does have a fair bit of trust of your PDS built in. But that's inevitable for decent UX—imo the crypto craze proved that basically no one wants to (or can) hold their own keys day-to-day. If you want to have a cryptographic protocol that the average person can use, some amount of trust is necessary. The AT Protocol artfully threads the needle and finds a good compromise that is a (large) improvement over the status quo, in my opinion.
Maybe your point of view is outdated?
People love to reinvent the wheel and claim it's a whole new thing. No ideas on the web have really been innovative since the bubble popped. The innovation has all been on delivery and execution (not wanting to discount any of that).
Secondly, signing every message wouldn’t be impractical at all, I don’t think. We’ve had the technology to do this for a long time and it’s very simple. What we don’t have is good key management. For average users, this would have to be something provided by their devices (phone or the Secure Enclave in your Mac or whatever) - managing keys and the web-of-trust shamozzle are the main reasons why encrypted email for everyone never took off.
So as long as both your source and destination support account transfers, you can usually switch and even seamlessly bring along most of your followers without them noticing.
No idea about your admin question. All bets are likely off with a bad admin. If you want actual cryptographically guaranteed communication, that doesn't exist in a usable form (except for Secure Scuttlebutt, and that's reeeeally stretching the "usable" part)
But none of your content goes with, and that's really terrible because it gives a lot of power to petty-kingdom jackass admins.
AT Protocol is heartening because maybe they've got a good solution to that.
(It's technically possible to edit the links in the posts you exported yourself before importing, but technically correct isn't the best kind of "correct")
You can pay for that service, but you have to administer the instance, and it’s not able to reuse the servers RAM for multiple domains; it’s not like email where spam management is built in.
Doesn't federation require admins of both instances to agree to federate?
If so, I could imagine it not scaling for larger instances that could receive 100+ federation requests a day.
Further, you would have to repeat this step each time to encountered a new instance you wished to participate on.
Is my understanding of the federation process correct or have I totally missed the mark? :D
My recommendation was much like MX records for email, so you can use a hosted server under your own identity.
The promise is that one can not only transfer the identity and all personal data across instances of a single service, but also across different services (imagine from mastodon to Lemmy).
Sadly, it never caught on.
This is why I want domains as identities to succeed. I want to own my handle on every platform, but I don’t want to self host.
Decoupling identity from social is a good idea but you can't just migrate the key storage to a single custodian entity. There'd need to be a multiple custodians to ensure the same power imbalances didn't reappear in a different form (e.g. Google owning everyone's logins).
All that is needed is a to create your local identity (e.g. like storing fingerprint biometrics on your laptop) and a clever way to sync between physical devices (e.g. through bluetooth).
We're in this weird situation where people don't want to be responsible for managing their own data/id, but can't trust others to do so for them.
I was toying with an idea/protocol where:
1. You add a TXT/CNAME that points to a trusted "authentication provider".
2. When you try and login to a website that supports the protocol, it checks the DNS record and redirects you to your provider.
3. You then "prove" that you own the domain to the provider - how this is done would be specific to each provider, but one possible method could be by providing a signed message that can be verified vs. a public key stored in a DNS record.
4. The provider redirects you back to the original website with a token.
5. Finally the original website consumes this token by sending it in a request to the provider. The response contains the domain as confirmation of the user's identity.
This approach removes the need for self-hosting as users can point and setup their names with third party providers.
Users can also trivially switch to a different/self-hosted provider by changing the CNAME.
Communities could also allow direct registration by hosting their own provider instance and pointing a wildcard subdomain at it: (i.e. *.users.ycombinator.com).
Users could then sign up to said provider using traditional email/password and claim a single subdomain: (i.e. tlonny.users.ycombinator.com)
Thoughts?
Though what you described is just a regular federated identity workflow, except autodiscovery through DNS (though that is already a thing for some)
Current spec drafts at <https://github.com/WebOfTrust/keri>.
I have been using Twitter more often lately as I really don't care about what they do to their APIs or whatever, and it's honestly improving lately.
Most of the people I followed before were clearly left wing. After Musk takeover a lot of them left the platform. Plus the algorithm now pushed more right-wing content in the home page. I wouldn't mind if it were real people discussing valid talking points. The problems are 1) they are all coming from blue check mark accounts, 2) most of it are clearly misinformation, and 3) you can tell most of these tweets and replies are troll and bot accounts. It's just annoying.
Musk boosting his own tweets in the feed was annoying. Had to unfollow him.
I use Twitter via mobile website, and it breaks more frequently than before.
Overall, it's become like Musk's other product, Tesla. It over promises and under delivers. It's not reliable anymore. As a 2x Toyota owner I cannot stand products that are results of crappy engineering. So there you go.
Edit: one more thing: it used to be that I could go to the trending hashtags and get latest news in a second. It's not the case anymore. Case in point: yesterday France was trending. I saw the tweets and got the impression that some member of minority community has committed mass stabbing or rape again. Because the Twitter results were brigaded by right wing blue check mark accounts spewing anti immigration propaganda. It was not until I read a BBC article that I realized what happened was complete opposite, police executed an immigrant at a traffic stop.
Twitter has lost almost all of its core values under Musk. It's just sad.
The majority use-case requires centralization, which is subject to the network effects that constitute 95% of Twitter's value. Great that it works for you and some others, but it cannot work for most.
Totally! The internet was better before pay walls / auth walls.
I get why we're here today, and I get that the last phase was just about acquiring an audience, and getting people hooked, and now this part of getting everyone to pay was always the plan, but this part really does suck.
Just feels like all the services lined up to start shitting on users at the same time. Netflix, YouTube, Reddit, Twitter, NYT (and all the newspapers really)... you can't watch Amazon Prime without having 300 "buy now" buttons in your face.
This phase of capitalism sucks.
Well, looks like I was only partially right. Most access is still "free", but at the cost of enshittification.
One would almost be tempted to ask for a cheque in the mail like in the old days.
But I guess for publishers it never produced enough revenue to be worth maintaining the integration.
The database has doomed us all! For it is the Seed of all Evil in the hands of the Wicked! ...and arguably the Good, but misguided!
Reddit's been circling the drain for years, but today is the day it truly crossed a line for millions of people at once (They killed Apollo and RIF).
Twitter's been getting rapidly worse since Musk bought it, but this is another red line.
We're not even allowed to talk about the influence of bots and astroturfers here. A popular post today was flagged within an hour, just for pointing out that a site claims to sell upvotes on HN.
I don't like it when the tech giants get in sync like this...
[1]: https://news.ycombinator.com/item?id=36526827
[2]: https://web.archive.org/web/20221215015113/https://pastebin....
- Reddit API lockdown
- Twitter account wall
- YouTube increases aggressiveness against Adblock
This is on top of similar behavior from other sites that happened long ago. Facebook, Instagram, LinkedIn, etc.
But drastic changes across platforms within 24 hours is unprecedented. This does not bode well.
The internet is fine. The highly centralized businesses that built themselves on the technology are meeting their inevitable end.
How would smaller companies do better?
For every gmail, how many independent email providers do you think exist out there?
Why would you choose to not pay for the big tech products and then pay for a smaller version?
Smaller companies and ISVs, on the other hand, will be better off if their market is commoditized. They won't have to spend so much to compete in R&D, they just need to find each the best way to serve their (comparatively) small customer base.
If https://en.wikipedia.org/wiki/Timeline_of_international_trad... is right, the necessary technology for international trade was not available before 4000 BCE (horse riding ~3700 BCE, long-distance sea travel ~2000 BCE).
One is that centralized _business_ has certainly not been the primary organizational mode. You can talk about centralized _government_ (of whatever variety you'd like), but the distinction there is that the centralized government had some sense of itself in an ecosystem - a citizenry, a land, a future, etc. - and businesses do not.
The second is that centralized entities sure wrote a bunch of stuff _down_, but it's hard to say they were the primary organizational mode for any but the last 50-200 years - the reach of serious centralized bureaucracies has only really begun to match their propaganda in the industrial and now computer age. Until extremely recently, the actual effective reach of a centralized bureaucracy was a day's horseback ride - control degrades rapidly as one leaves the core. "Heaven is high and the emperor is far away", as the saying goes.
Edit: With regards to the second note, "Against the Grain" by James C. Scott is a solid read.
If you own a company and want it to last through the ages and you aren't literally the only guy in town, never become number one. Aim high, but don't hit the top.
The same can be said for countries and practically any organization or group. You stay an underdog if you do not want to ever fail.
That leads to complacency, corruption, and delusion, ultimately leading to failure.
Everyone and everything from mundane individuals to megacorporations and empires have all fallen from grace once they became top dog. No exceptions.
If you want to last, aim high but don't become top dog.
For most of human history, most businesses were small ma and pa shops operated by a few local people. These days large business chains are the norm. You could say that centralized big business killed the decentralized small ma and pa shops.
I suspect that was the intent.
I suspect the later. We’ve seen billionaires do it countless times.
[0] https://www.independent.co.uk/life-style/health-and-families...
I used the 'redirect to nitter' Firefox extension and Android app but it got quite unreliable and nobody else that I know uses nitter at all. I think Nitter users would be a tiny, perhaps even immeasurable minority compared to casual readers that now are incentivized to either log in or fuck off (but... FOMO).
Of anything I suspect Elon saw what happened woth reddit and had a "wait, we should do that!" Moment.
On every large platform, this hasn't been the majority for a long time.
Hah, this might mean that we'd be one step closer to the dead internet (conspiracy) theory actually being reality: https://en.m.wikipedia.org/wiki/Dead_Internet_theory
What an interesting concept.
Publicly free content from users devolves to garbage content. I think the Chat GPT effect is they're realizing its easier for companies/entities to generate garbage that is at or exceeding the intelligence of comments by actual people (a low bar). Sure there are pockets of usefulness but this is tiny amongst the firehose of garbage.
If all that is publicly available on social platforms is just garbage nonsense, people will just stop going if any barrier is thrown in front of them. The internet as a technology stack is fine. This is how social media dies (hopefully).
A potential saving grace: I bet within a year or so it will be easy to self-host LLMs that are easy to fine tune and run. Then there will be a few open source tools that you can use yourself, privately, to capture your level of interest while reading, and periodically make a reader/summarize/filter agent.
This is not scary if people can fairly easily run it all themselves, keeping their data private. It would help wade through crap, and there is some irony in using LLMs to have personal filtering and summarization.
Compared to future versions of ChatGPT, Bard, etc. the models that individuals can self host will be much weaker, but I think that they will be strong enough for personal agents and eventually be cheap to run, and affordable to fine tune.
The clock is ticking until you’ll need to decide which internet package you prefer, based on content providers and limited by region.
At least Yahoo will probably come free with the base package.
Those sites never fit my definition of open anyway (free, permissively licensed technology and content). The ones that do are smaller, aren't a monoculture and seem to be pretty untroubled so far. No one wants to scrape the little Mastodon or Lemmy instances or other small community sites I pay attention to.
Big deal if something is a threat to the social media business. The social media business is a cancer which should be destroyed anyway. Go outside and engage in real socializing instead of the depression-spawning, teen-girl-murdering version peddled by Zuckerberg and Musk. It's much better and once you change your habits you'll never look back. Maybe it has something to do with all the vitamin D you get from being outside.
We've BEEN through federated platforms before. We've even been through PROTOCOLS before. They're all horrible. The successor to any platform that currently exists will have slight improvements to what already exists, and that's IF they're able to do so.
I don't have a dog in this fight, but I do have over 30 years of being around social media platforms on the internet.
Remember, this is a VC firm that runs the joint. We're paying with our data, and that gives them early/first mover advantage.
If Twitter locks out more readers, people will stop posting and move elsewhere. If Reddit ejects the mods that made it successful, communities will evolve elsewhere.
The forest will regrow and different paths will form, routing around the dying patches.
The only time I run into this Fuck You pattern is Instagram refusing to let you replay a video.
They aren't trying to delight you, or make the best product for you.
They're trying to suck as much money out of you as possible. If that means making a worse product, you're gonna get a worse product.
Until this gets beat into the average person's head, we'll keep falling for companies that we think will be different.
Makes this all really easy. It is a big world out there with lots to do, build, see, laugh, love, play.
Fuck em. I just do not need it. Whatever it is, the vast majority of the time.
It may still be rough around the edges, but to me it feels like the spirit of the old phpBB forums combined with almost 20 years of lessons learned from Reddit.
But to have something pay for it self. I'd rather not lose money on it
I was on a forum, and the kind of language and even topics for discussion were severely restricted, partly because of advertising.
Matrix does better as the rooms are decentalized so can continue even if the creating server drops offline. But user accounts are still only federated.
https://old.reddit.com/r/Save3rdPartyApps/comments/14k67qt/l...
Highly sus
https://github.com/LemmyNet/lemmy/issues/622
I would not bed with these fellows myself.
Maybe they were abrasive in initially fighting the request to make technical changes to the slur filter, but hey when you ask for free enhancements to open source code you either do the work and provide a pull request or be prepared to be told no.
I empathize with their concern about becoming another Voat or Gab. They want federation but they don’t want a Wild West.
So the word above word is used in lyrics of a music genre with predominantly black musicians. In addition to saying we don't want our software to be used by racists, they also say "we don't want our Software to be used to discuss certain kinds of black music" (arguably a racist stance just by itself). Talk about unintended side effects.
if automation is chosen there will absolutely be situations where perfection is impossible. if human’s unparalleled ability to see nuance is chosen then the cost scales along with the amount of information.
the fact is, if we want a community and we want to keep signal above noise, we will need some form of removal of spam, child porn, racism, etc…
automatic tools can’t nuance as well as humans.
then human mods start nuancing and someone will point at stuff and call it biased.
It did not seem to me a politicized discussion but a technical issue with filtering using hardcoded blacklists that are just too prone to the Scunthorpe Problem. Perhaps because too many people in the USA despise the mere existence of other languages :)
As far as the “popular” list, which is placed below the recommended list, lemmy.ml doesn’t have any special privileges there. It just happens to be the most popular. If something becomes more popular it will go on the top.
The problem with Lemmy is that one gets sent to some place like https://github.com/maltfield/awesome-lemmy-instances, is immediately confronted with a ton of weird links like "butts.international" and "badblocks.rocks", what even is this? And about 100000 other servers just named "lemmy", "notlemmy", and "lemmy1". So you click a few at random optimistically, then get hit with login page, or a server error, or an apparently empty test-server. You begin to think you're being pranked, like am I supposed to brute-force click like 50 things to find something that's not a joke? Maybe you go to https://join-lemmy.org/ and it says "After you create an account, you can find communities", so great, it's inaccessible anonymously, the same as twitter. You go to https://lemmymap.feddit.de/ and after 15m of page-loading get a hilariously useless cyberpunk-looking word-soup where you can't click any links, much less search for topics/communities (btw there are 2068431 running instances and somehow butts.international is still front and center in my cyberpunk view)
Finally, by ignoring recommended tooling and just using google-search I found a community relevant to my interests, but it has pretty bad content and a whopping 1 user/day. Another google search trying to find a certain topic, I find one, but it has only 3 total comments, and I could not tell what month/year the posts were added.
So, clearly I don't really know what I'm doing here, but this stuff is ridiculous. As long as we're crawling 2068431 instances why don't we look at the communities it hosts and the volume/recency of traffic? At least filter totally empty stuff and/or make it easier to get all the test instances in a sandbox! Discoverability is so bad that I can barely get to the point where I'm considering usability / content.
Between account walls and search’s indexing problem, it’s become very hard to find small to mid sized active communities on your own. In fact this problem seems to be something people are trying to solve in Reddit communities via related subreddits on the sidebar.
So having gone from using search engine’s to crawl for relevant content that was out there, people are now creating content specifically to end up in Google’s search results - destroying the value search once had. Indicated by what people have done on Reddit, and these discussions about finding alternatives, it seems we are well on our way back to webrings. I welcome this.
This is not true. Communities can be browsed without login on their home instances.
For the short term at least, since discoverability is so broken, I think those who want to advocate for Lemmy will be better served by just linking to content or curating indexes of active communities. It's not that useful to anyone if the focus is always about pretending everything is fine, or presenting prospective users with totally useless machine-generated indexes where we cannot tell the test-servers from production.
And then it will magically open back up. (More press)
And that's the story of pretty much every one of the outrageous/bold/brilliant/terrible strategic decisions you've heard of since the Twitter takeover.
What's most amazing is that it works everytime, I'm surprised there isn't an Onion copy/paste article about this each time.
My Twitter usage is down a lot since they killed off 3rd party clients.
Every user hostile move they've made basically halves my usage.
I used to spend way too much time on Twitter, now it's 10 minutes a week and I don't feel like I miss much.
Just a boring echochamber.
Its the twitter equivalent of youtube thumbnail face with 10 million $$$$$$$ in the title.
He is killing engagement and ad income, has to slow the bleed, and is now looking under every rock for some gains, GOTO 1.
I've been meaning to make a Twitter account since I follow a few accounts for some games that I play.
Between Twitter not being owned by a loon anymore and finally a reason to overcome my laziness, why not?
And yes, I know I'm playing right into Twitter's hand. Whatever, I actually like Musk anyway (an unpopular opinion that will no doubt get me flagged around these parts).
But that's clearly not sustainable, it's inevitable they'll pivot to monetizing the service - and that's clearly going to make it less attractive than the heavily-subsidized version people have gotten used to.
I guess the ideas is to lock the users in to the degree that the increasing monetization is put up with, and slowly enough there's no "sticker shock" of a previously-free service suddenly having a price.
I'm constantly amazed people are surprised by this, isn't it obvious that tying to a loss-leader service isn't sustainable?
Instead with their VC funding they spent $$$ to introduce things like NFTs and TikTok scrolling, which nobody asked for, but it burned through a lot of cash.
This is actually the most mood-lightening comment I've seen on this matter. A return to the internet of 15 years ago doesn't sound so bad to me. (Of course, the mood darkens again when I remember that it won't really be that, because, e.g., that internet of 15 years ago couldn't handle the bot-spam of today.)
I'm at least glad that people will start looking into the alternatives.
The next lesson they need to learn is TANSTAAFL. Nothing good can come with an internet where publishers are paid with eyeballs. We need to rescue the idea that "voting with your wallet" is the best and fairest way to have quality content.
people been saying this for a decade +
you'll be back
What sucks is people flocking like sheeps to centralized platforms that have not their best interests at heart.
Decentralization has been the answer from day 1 but people dont understand shit and make the same mistake over and over again
i havent' logged into twitter since forever but still have accounts
i need to delete them
As long as the site is growing, it doesn't have to be profitable. But when the music stops...
Years ago I was brought to twitter compared to Facebook exactly because you could read without being logged in. After a few years it had become a hellish place with lots of flames and arguments, but it still had some value. It became clear that my engagement was mainly to discuss with random people about things knowing they would never change their minds (neither would I) on things like Covid vaccines. It was a huge waste of time, but I found it out that surfing it as non logged would amore allow me to read without being able to reply to the most stupid comments. Some sort of read only Twitter. Now that it has gone, Twitter has irrelevant.
Time to go outside and forget that the online world exists.
If we as a society decided to channel these efforts into building infrastructure improvements and homes, I think this would help a lot more. I understand the biggest problems in that respect are legal and cultural, but I can’t help but feel people have tried nothing and are all out of ideas.
(More precisely: we're acutely aware that users hate change, and since we do too, it's kind of an easy call.)
Also, the value of HN to YC consists of the community and keeping the community happy is therefore a must.
(Did I say happy? More precisely: as happy as possible under the circumstances)
That's actually a nice way of putting it! Better than the "fix" version, because this is clear on the consequences.
I feel like it might be applied to everything from OS UI design (Windows 11), web platform redesigns (Google's icons) to whatever is going on with social media (the silly enshittification term describes this) and many other things.
Let's say that you are the proud owner of a goose that lays golden eggs. "fixing" would be switching it to a different feed that might make it more productive, or it might make it sick. But this year the trend is to give it a few good kicks to see if that helps.
https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que...
(I've a slightly more updated set locally, can share those if requested.)
HN could be hit with such a large load too, since we have pretty good and lengthy discussion here, good for AI data training.
Would you believe it a valid tool to keep the community happy?
Since I doubt we would be happy if we can't access the site because some AI decided this was the time to scrape, either.
If the world around HN (including its community) changes, stasis can damage or kill it as well.
Specifically regarding the issue of the original posting:
- HN is already an important data source for large language model training. [1]
- To the best of my knowledge there is no freely downloadable and current data dump of HN. [2]
- The HN-API does not offer all the data that scraping can get. For example, if a post had ever hit the front page or the highest front page position reached, is an interesting data point that is missing.
- The Algolia-HN-API has the same limitations.
In my opinion this will lead to increased usage of the API and increased scraping which all costs money. HN might be forced to find a solution for this.
[1] For example, the RefinedWeb paper lists HN as one of only 12 websites that were excluded. From what I understand, it was excluded because it went into the final dataset unvetted. RefinedWeb was used for the Falcon model.
https://arxiv.org/pdf/2306.01116.pdf
[2] The closest thing is probably the Google BigQuery "bigquery-public-data.hacker_news" dataset. It claims to be updated daily, but really is from late September 2022. Also I could not find the download link which other data sets offered on BigQuery have. Does anyone know if I can download the complete thing anyhow?
https://console.cloud.google.com/bigquery?p=bigquery-public-...
The one thing I see as a future issue is that people are starting to post comments that clearly look like they were manufactured by ChatGPT and friends. Or that could just be the way some people talk and I've spent too long with ChatGPT now and start to smell it everywhere.
I've had that happen even under manual browsing (when logged out). My front-page analytics project hit that limit quickly (within about 30 requests, probably less). Adding in a reasonable delay got around that.
Keep in mind that a lot of Web infrastructure tends over time to operate just at the edge of stability, as capacity costs money.
https://news.ycombinator.com/item?id=22930391
[dang]: "Push notifications seem to jack up the nervous system in a way that's good for engagement but not necessarily for users..."
It’s not that I’m opposed to change really. I love good ideas, and being surprised by new and unfamiliar things is usually a joy. Communication via text is hard to improve upon though, and I’m not convinced any major social media platforms have found ways to improve this in any meaningful ways.
I want to read interesting things and discuss them with interesting people. This is hard on most platforms. HN makes it easier than everything else I use.
I believe this eternal loop of change is a trap that impatient people force themselves in. Instead of accustoming and learning the current state, they rush into another one with a change that has unclear implications. As a result, they never get where they are and lose any track of where they were or where they’re heading at. Their only comfort can be found in a constant change.
These few sites are like home to me. One of them I visit with years-long pauses and every time I return it’s the same user experience. That’s invaluable.
I thought that was an odd change for HN. After all, the majority of the value still accrues to the HN owners. In fact, user's that are prepared to pay for an app to access a free website largely comprised of adverts, are typically more valuable than the rest. Those users have money to burn and skin in the game!
So yes, occasionally some new pattern or trend will emerge, but HN adapts to those fairly quickly.
I've been sort of live-blogging the experience on the Fediverse: <https://toot.cat/@dredmorbius/tagged/HackerNewsAnalytics>, as well as in some of my HN posts.
My current tack involves looking at sites (as reported in parentheses at the end of each HN front-page post title) and classifying those. With slightly more than 30% of sites categorised, I can classify about 65% of all HN posts.
For the full dataset (17 years), that's roughly:
1 63913 35.73% UNCLASSIFIED
2 22589 12.63% blog
3 15112 8.45% general news
4 13823 7.73% tech news
5 12851 7.18% programming
6 8622 4.82% corporate comm.
7 8459 4.73% academic / science
8 7294 4.08% n/a
9 5324 2.98% business news
10 3803 2.13% general interest
11 2151 1.20% social media
12 2074 1.16% software
13 1613 0.90% technology
14 1463 0.82% video
15 1144 0.64% general info (wiki)
16 1009 0.56% government
17 724 0.40% misc documents
18 720 0.40% law
19 702 0.39% tech discussion
20 620 0.35% science news
Tons of caveats: this depends heavily on how I classify individual sites, a given site's stories might well be technical, social, or political, etc., etc.The breakdown-by-year analysis is in development, but if anything programming-specific content as increased in prevalence. Political discussion seems not to have (though it rose significantly ~2014). Cryptocurrency and blockchain-specific sites also peaked about that time (I suspect much of that discussion is now mainstream). General news has always been a huge portion of HN discussion, as have individual (and corporate) blogs.
Note again that this isn't about discussion and comments, or even the titles or article contents (I'm thinking of looking at those, it's ... a challenge for me).
But across nearly 200,000 front-page stories, on which nearly half of all HN discussion occurs (based on another API-based study looking at comprehensive posts), the overall trending seems at first blush to be pretty consistent and if anything improving over time.
(As with all preliminary results, I'm hoping I won't have to eat my words here. Though I'm reasonably confident in most of this.)
From the classifications above, the places you might find some that "suffering" would be in general news, genral interest, and social media categories. All but the first of those are single-digit percentages, and a lot of that general-news content is about technology, business, finance, and science, all of which would crowd out the sort of social and political issues which seem to generate strong feelings.
The "UNCLASSIFIED" sites are a wide mix, though most are probably a mix of blogs, corporate / organisational communications, and the like. The mean posts per site is 1.739951, so gains from additional site-categorisation are pretty slim. I have captured a lot of obvious patterns via regexes and string matches, so academic/science and major (or even minor) blogging and social media sites aren't a large fraction.
More recent discussion here: <https://news.ycombinator.com/item?id=36524001>
It's not practical to list all of them. But we can randomly sample. And large-sample statistics start to apply at about n=30, so let's just grab 30 of those sites at random using `sort -R | head -30`:
1 sfg.io
1 extroverteddeveloper.com
2 letmego.com
1 thestrad.com
2 bombmagazine.org
1 domlaut.com
1 bootstrap.io
1 jumpdriveair.com
2 desmos.com
1 leo32345.com
1 echopen.org
1 schd.ws
1 web3us.com
7 akkartik.name
1 bcardarella.com
1 cancerletter.com
1 platinumgames.com
1 industrytap.com
2 worldoftea.org
1 motion.ai
1 vectorly.io
2 enterprise.google.com
1 lift-heavy.com
1 davidpeter.me
1 panoye.com
3 thestrategybridge.org
2 fontsquirrel.com
1 kettunen.io
1 moogfoundation.org
2 elekslabs.com
That's a few foundations, a few blogs, a corporate site (enterprise.google.com), and something about tea, all with a small number of posts (1--7).I'm looking at some slightly larger samples (60--100) here on my own system, and can actually make some comparisons across samples (to see how much variance there is) which can give some more information on tuning what I would expect to find under the "UNCLASSIFIED" sites.
I've written about this a fair bit over the years if anyone wants more: https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so...
* ("good" here meaning "as good as possible under the circumstances")
But may be it will increase your costs a lot.
posts not being able to submit images is a huge part of what makes HN valuable (to me).
once HN needs to handle images, HN falls apart
I wish all the other sites on the Internet would wake up every morning, look at their TODO list and say "Nah, not today."
Replace http://twitter.com by http://traittor.net
http://twitter.com/elonmusk/status/1675187969420828672
to
http://traittor.net/elonmusk/status/1675187969420828672
Is normally bypassing no registered account and limit by day.
Unfortunately it does not work with all the tweets and especially the recent ones
Have fun :)
First Twitter API, then Reddit API, so today Apollo and many more Reddit clients shut down, and now Nitter. :-(
I'm happy Lemmy is kind of taking off. I think it's helped more than Mastodon because it's less realtime/feed focused and slower paced. It also doesn't require you to form a friend circle to benefit. Instead, the community is waiting for you already. You just sign up on an instance and add your communities. Done. This helped me a lot, together with sites like https://sub.rehab
It's one of the nice things about AP is that whatever interface you like best Twitter/Reddit/RSS can get you access to the same content.
Is Lemmy named after Lemmy?
I ended up starting at programming.dev because someone on HN mentioned it and it at least seemed to have a focus and also wasn't a ghost town. And that was pretty good but I've also joined beehaw (takes some time) because I like its size and decorum and generically would choose to get on their side of a defederation. And after I'm starting to understand how this whole activity pub and defederation and federation works, I really am optimistic about it.
I think somebody needs to build something that's a crossover between GitHub pages and activitypub that sort of behaves like discus and integrates with Lemmy/kbin/mastodon. So that blog writers can have comments at their own sites again and they can integrate together to grow organically. I haven't quite pieced it all together, but I sort of see that could grow as a replacement for what we lost with Google Reader and the loss of blog commenting communities.
Agreed. I think the barrier would be lower if I knew I could migrate my identity to another instance if the first one became sketchy or shut down or de-federated.
Instead AFAICT I have to choose not just what community to join and where the content will initially live, but also which of these random groups to trust with my identity indefinitely going forward.
Like SSH keys, where you manage your own identity and then share a public key to each instance that identifies you to that instance.
Like an identity client you could self manage if you wanted to. Make it optional, portable, and transferable. So you can choose to let a server host manage your identity, or migrate to a self managed identity.
If I had more time on my hands….
Conceptually a planned migrations should have a period of concurrent access to both the old and new accounts and it should be easy to publish a handshake to confirm the migration to followers to update contact info. That's my thought anyway. Something like keybase (does that still exist?) could also be used for similar sorts of proofs.
The way people used to handle this on Reddit is people would send a message from their old account saying "hey XYZ is my new account" and that seems sufficient.
Of course, this goes for any major interest category but it just hit me the hardest so far to realize this.
Another cool development would be a science-oriented Lemmy instance with lots of special purpose sciency stuff.
Viewed like this, the sky is the limit for Lemmy and it could have potential to grow a lot!
Something like this?
https://mander.xyz/communities
> An instance dedicated to nature and science. > The main focus of this instance is the natural sciences, and the scope encompasses all of the STEM fields.
It's hard to map this onto my framework of "reasonable paranoia". Even while I felt uncomfortable about it, it never occurred to me that Twitter would actually cut off access. Now here we are.
Signing up for a Twitter account is free and can be done by literally anyone in the organization.
The answer is because Twitter is easy and free. Self-publishing is neither.
i think the real question is when, not if, we end up with public infrastructure.
The government built the Internet. They can host a website.
Hint: it's not the second one
No change control there either.
and the only place to stumble onto information is where one spends the time - facebook, reddit, heck - even hackernews
It's not where I would look for information. But that assumes I know I should be looking.
Only to the extent allowed by news websites headers. Maybe this whole thing could have been solved by politicians understanding tech a little better?
How often do we hear, on HN and elsewhere, people complain about laws they don't like wish politicians better understood what they are legislating? This implies the existence of a technological solution, set by technocrats who "understand" better. Whoever asks the question suggests politicians don't know better, offering that explanation without proof.
Also, for the Free Market enthusiasts out there, why hasn't the market solved this problem? What forces are preventing all parties from working out a technical and economic solution?
This is like mandating car makers to pay whip makers.
> Also, for the Free Market enthusiasts out there, why hasn't the market solved this problem?
Why are assuming the market hasn't solved it? And that too without proof?
> Whoever asks the question suggests politicians don't know better, offering that explanation without proof.
It is easy. There is zero discussion of technology beyond sharing of links in the bill.
Goodbye Meta and Google and other greedy foreign exploiters and don't let the door hit you on the way out.
So they can never make mistakes and be blame free in any scenario?
> refusing to extract their wealth if it can not be done for free, and you're entirely correct.
Yes, now they can keep all their wealth with them. Why are they complaining? Why fret over this if extraction is being stopped by these companies shutting down link sharing?
> Goodbye Meta and Google and other greedy foreign exploiters and don't let the door hit you on the way out.
Yes, it is the foreign companies who are greedy and not rent seeking mega news corps who want $$$ for linking to them lol.
There are no large media companies in Canada, just independent journalism pure in heart and no commercial intent. /s
> greedy foreign exploiters
So xenophobia is a cause for this action?
The law was about news media, Facebook shutdown a bunch of unrelated govt services pages in retaliation.
Facebook definitely deserves a lot of blame. Likely in this case too. I'm pro capitalist as much as the next guy but laws are by the people, if a business wants to extract money from our community it needs to play by our rules.
Facebook cutting access is like a badly raised toddler throwing a tantrum.
> I'm pro capitalist as much as the next guy but laws are by the people, if a business wants to extract money from our community it needs to play by our rules.
Wait, if FB was extracting money, wouldn't FB shutting down linking cause more money to flow into news orgs? Shouldn't you celebrate this as it will reduce the "extraction"?
Why fret over this if extraction is being stopped?
If news orgs truly believe in extraction, they would be celebrating this shutdown in the streets.
> Facebook cutting access is like a badly raised toddler throwing a tantrum.
In the free world, we are allowed to choose our actions to some extent.
It is very clear that it is due to them not wanting to comply with bill C-18.
Reported on by every major news outlet that I’ve seen
https://www.theguardian.com/media-network/2014/dec/12/google...
It remained closed until 2022 when the law was changed and newspapers were allowed to negotiate directly with Google individually.
https://www.reuters.com/technology/google-news-re-opens-spai...
> The move comes in reaction to the federal government's Online News Act, Bill C-18, which would require the tech giant to pay Canadian media companies for linking to or otherwise repurposing their content online
How is Twitter any different than FOX News or CNN?
There's a big difference.
Governmental and public institutions should not be relying on Twitter for their communications. It's unprofessional at the very least.
You wouldn't put quotes around that word for cell-phone companies using licensed bandwidth, or airlines using public airspace, would you?
And they are not private. The public is legally entitled to receive anything broadcast over the radio spectrum. And as I and another poster have already pointed out, there are government licensing and carriage requirements involved with TV and radio broadcasters.
Local water systems in Guam are privatized, much to the chagrin of local activists.
Elsewhere, there's an article on the guardian just like yesterday about the impact of privatization on either the UK's or some locale inside the UK, for their water usage. Basically the water department is now the highest debt entity in all water departments and it's the one that's privatized.
Privatizing public data is a shortsighted thoughtless approach to public communication.
So Play stupid games went stupid prizes.
Coming from the UK I don’t want competition in basic utilities like there is there. I don’t want to have to shop around for the best energy price every year, or to be at the mercy of the free market on pricing for the most basic of necessities like power and water.
It's the hysterical marching cry of young naive people railing against the ever-increasingly vague boogeyman of neoliberalism but the actual evidence doesn't back it up.
You talk about ideology over logic while demonstrating that exact same thing.
> The AER report contains no consistent correlation between higher bills and privatisation.
> The ABS index of electricity prices across Australia, showing movement of electricity prices over time, also doesn't demonstrate a link between privatisation and price rises.
> Whether comparing electricity bills, prices or the relative price index of electricity in each state, there is no consistent link between privatisation and what consumers pay for their electricity.
> Experts say the biggest influences on what people pay for electricity are costs of transmission and distribution. They say these costs have risen in recent years irrespective of whether the owners of the transmission and distribution networks are privatised.
https://www.abc.net.au/news/2015-03-25/fact-check-does-priva...
Will you entrust your well-being to CEOs and boards whose sole priority is relentless growth and maximizing profits, regardless of the consequences?
Laws can only do so much to prevent vital resources from becoming unaffordable for the very people who deserve them. These companies, with no regard for your, the citizen's, vote or input, prioritize profits above all else. To them, it's just another business move, leaving the consequences behind as they move on to the next venture.
You can argue and consider other ideas, but the higher risk in the private side is always there.
And they are easily reigned in by the Australian Electricity Regulator, while governments owned businesses are not.
We recently had a price cap put on coal and gas prices, private companies had to eat a massive loss, government owned generators huffed and puffed until they got billion dollar bailouts from the rest of the nation for their coal guzzling power plants that should have been shut down years ago had they not been taxpayer liabilities.
There's simply no evidence of these things driving up prices. We have a huge government owned pumped hydro power storage project underway that has blown out from $2b to $10b in the spacce of a few years, if this was privately owned it would have gone bankrupt, instead more and more cash from electricity users will eventually have to be paid to fund it. It's costing 5x more than simply putting grid scale batteries in the cities and the government has sunk cost fallacy that it can't walk away from.
This is the problem with government ownership of electricity assets in Australia, politics gets completely in the way of good decision making and projects that should have failed long ago.
As I said before, where's the evidence, because all I see is ideology.
This is what happens when governments run electricity projects:
> He assured the electorate it would cost $2 billion and be up and running by 2021.
> By April 2019, a contract for part of the project was signed for $5.1 billion — and that doesn't include transmission costs, which will cost billions more.
> Who will actually pay for transmission is still being decided.
> "Someone's going to pay for it," Snowy Hydro CEO Paul Broad told 7.30.
> "The taxpayers will pay for it through your taxes, or you pay for it through your bills.
The current cost is $10b and the project has blown out to the end of the decade. No one is admitting how far they've dug or the status of meeting targets. The transmission bill is still up in the air but will likely be an additional $5b added onto household bills.
https://www.abc.net.au/news/2019-10-14/snowy-hydro-2.0-expen...
Also, the LNP have proved time and time again they can't build any public infrastructure.
Their Inland Rail project is another total disaster, a project wasting many more tens of billions and achieved absolutely nothing.
And how can anyone forget the failure that was the LNP re-design of the National Broadband; a design that moved away from fiber optic to instead use copper wire.
Put up your "logic" rather your ideology if you want to convince people otherwise. Downvotes don't count sorry :)
Prices are the cheapest in Victoria, with full privatisation of the network. Prices are also most expensive in South Australia, with nearly full privatisation of the network. No rational person can look at that and proclaim there is a correlation.
We have states with nearly all generation, distribution and transmission being government owned, we also have states completely out of the business. This should be a simple slam dunk to people who loudly make these easily proven claims and yet they never have any proof.
Prices are rising uniformly across the board, including in the heavily government owned states, which coincidently are the worst at rolling out renewables because they are protecting their fossil fuel golden eggs at the expense of the environment.
https://theconversation.com/myths-not-facts-muddy-the-electr...
While interesting time capsules one wonders whether Lynne Chester holds the same opinions today or has updated the prices tables 2007-2014 with more current data.
Personally I'd look less at the month to month prices and more at the decade on decade projections .. what are the winning long term strategies for cost effective power generation with the "lowest bad for climate" emmissions totals.
Is it a coincidence that Western Australia (WA), a state with the most government regulation in regard to LNG exports also has the lowest LNG prices by a long margin?
In general WA, has the most regulated energy sector with the highest level of public ownership in electricity production and transmission, and by strange coincidence it also has the cheapest gas, cheapest coal and lowest electricity prices.
The profit motive.
I don't actually think the Australian grid is a good example of privatisation failure. There's little proof that we have meaningfully lost anything aside from the marching cry of the left.
Did we gain anything? I ask in earnest, I don't know.
I can guess that the cost of running, maintenance, etc are now "not a cost of the state", but the cost doesn't just disappear, it was always on the consumer via tax or bills (or tax and bills, hopefully in a way that sums to the same cost!).
As it pertains to power, People tend to promote government as a means to make it cheaper - but this is sort of a fallacy - the only way the government can do it cheaper is if it's operationally more efficient to a significant degree. I don't see how a government achieves this. Private energy generators are usually not very profitable - the people making money were commodity producers, recently, not electricity generators.
It’s not just cheaper, it takes the privatised profits off the table and instead feeds them back.
It’s not magic, it’s pretty good sense though, and it serves the people of the state well.
IMO this should be illegal.
Google is more like a malaria carrying species of mosquito. Little bites that you barely notice, but in aggregate are actually much more dangerous.
I know the metaphor is a bit stretched but even in the late 90’d it was clear what Microsoft did. Google has gone from “don’t be evil” to “this is the definition of open source” to …
Of course my point was that the mosquito, whose bite is far less “annoying”, is actually #1.
[0] https://www.sciencealert.com/what-are-the-worlds-15-deadlies...
These same government spies run an international network of torture centers.
And I can guarantee you they are going to bring in high priced consultants to do the work because no government or board of education is going to ever pay software developers their market value and have them on their payroll.
Open source is more secure is just as much of a fallacy.
In a given month, as a consultant I’m working with clients that use Slack, Teams, Google Meet, Zoom, and if I’m initiating a meeting, Amazon Chime.
I wish, the US government, was like that full of decisive, well intentioned, intelligent boot strapped innovators but the motiff of government projects are things like:
One bathroom stall at build costing millions and years behind schedule
City trash bins costing hundreds of thousands
Airplane trash bins costing tens of thousands
And the list certainly Curtis go on those are just the recent ones I’ve read about.
I’m watching a government software replacement worth tens of billions and seeing the attempt to integrate is painful. They sure aren’t agile.
I don’t work in tech.
How are they ever going to be able to compete for talent?
You'd need to convince people to side-load an app from a government website, or install a government app store.
But why? It’s always been that way. Before Twitter, it was private radio stations, private TV stations or private newspapers.
In the end it’s still private companies (with the exception of public broadcasters or websites).
You were able to access those news for free. Without any account. Once you had a radio or a tv, it was free and accessible for everyone.
Not the same for Facebook or Twitter. Even if technically free, you can be banned, or have your account deactivated because you didn’t give away your phone number as “security measure”.
It's not obtuse, there are a ton of similarities.
Also is for example BBC private company?
You can listen to TV and radio stations for free without an account or subscription and they can't cancel you (unlike FB/Twitter/etc).
Newspapers tecnically you had to buy, but reading the headlines at the newstand was free and you could always go to the library of coffee shop to read the whole thing for free.
Weirdly, Twitter had started becoming so unreliable for me for several months prior (frequently not loading, video rarely working) that my click-through rate on Twitter links was already diminishing. But looks like it’s 0% from here on out.
I haven’t missed it.
I think it speaks a lot about the _illusion of value_ these networks provide vs actual value.
Yes it's enormous and popular now, but that's in spite of musks product vision.
https://chrome.google.com/webstore/detail/hide-blue-checks/a...
We are are not that sophisticated of animals. I mean we're not as dumb as dinosaurs were where they forgot about prey if the prey turned the corner out of their sight, but we're not as smart as we want to believe we are. So I think Twitter is Entertainment and Should not be relied upon.
"Temporary emergency measure. We were getting data pillaged so much that it was degrading service for normal users!"
Then I realized... hello there! (:
I was going to host the screenshot on imgur, but I'm not sure if we can trust them anymore...
"Please respond to the strongest plausible interpretation of what someone says, not a weaker one that's easier to criticize. Assume good faith."
I'm not saying you owe CEO billionaires or billionaire CEOs better, but you owe this community better if you're posting here. If you'd please review and follow the site guidelines, we'd appreciate it: https://news.ycombinator.com/newsguidelines.html.
As far as snark, I see several other examples of that—also intended for Musk—in this thread. They don't strike me as either offensive or particularly constructive, so it's not clear to me why my comment was called out here. (Especially considering that there are a few other comments that definitely go beyond the acceptable levels of user-to-user snark as I understand them.)
I can avoid sarcastic comments about billionaires in the future, if that's a problem. If the issue was snark directed at another user, that wasn't my intention.
I'll also say that the "snark" rule you cited, while well-intentioned, seems very broad and selectively applied here.
It does not say “don’t be snarky unless scarasm is directed at a billionaire because then it’s ok because they have a lot of money and power, so we will allow it”.
You would then need to define some amount of money that would put someone in then “can be flamed” category.
The rule is not applied selectively here; it is applied to everyone, Musk included.
I'd say what's not clear to me is what to avoid in the future. I'm not trying to be difficult here—and dang is one guy dealing with the internet version of a city, to be sure—but I see sarcasm all the time on HN. The really toxic, demeaning stuff, sure, that has to go. In this case, it never even crossed my mind that what I said would be interpreted as targeting the person I was replying to. (While I wouldn't have flagged it, your snarky response, by contrast, was pretty clearly targeting me.)
Looking over the thread—and HN in general—there are no end of snarky posts, including yours, and especially in regards to wealthy tech guys like Musk. The vast majority of them are permitted. That's what I mean by "selectively." Going by your interpretation, no snark would be welcome at all; if that isn't the case, which I didn't have the impression it was, then what was it about my post that warranted a response more than the others?
Genuine question. I can observe consistent rules, but I'm not seeing consistent application of this one.
The reason I saw your comment rather than the other ones is that it was heavily upvoted and right near the top of the page—and that's just the problem: snarky, shallow comments attract upvotes, which causes them to occupy prime real estate, crowding out better discussion, and that distorts the character of the thread and ultimately of the site itself. This is one of the biggest problems HN faces, if not the biggest.
You can say that this problem is caused more by upvotes than by comments, and I agree - but we can't address the problem at the upvote level (at least not publicly), and anyway if the flypaper weren't hung in the first place, the flies wouldn't have thronged to it.
It's impossible not to "selectively apply" the rules, the same way that not every speeder gets a speeding ticket, and of course when you've seen other people speeding worse than you (which they invariably do), it feels unfair that you're the one who gets pulled over. The main things to realize are (1) it's nothing personal; (2) the randomness evens out in the long run; and (3) the only way to keep HN going in a good way is for enough commenters to understand this and take up the work of following the site guidelines (or really, the intended spirit of the site) even when they see others not doing it. I hope this helps explain things a bit...
It's the same with reddit, the user is the product, which is why they don't want you using third party apps/frontends.
Secondly, that "data pillaging" was exposure, he just removed exposure from people's tweets to save money, again.
Good work Elon.
With all the public officials on Twitter (and FaceBook) publishing "public-facing" information, I'm surprised both are allowed to be/remain walled gardens.
To do my business taxes this year I had to DEMAND paper form acceptance, which was begrudgingly accepted once I went, in person, to the tax authorities [they want you to provide all sorts of tracking JUST TO FILE STATE TAXES, under the auspices of "two factor authentication"].
As a technosophisticate that intentionally avoids email and doesn't carry a cellphone... I weep for what my public-interfacing world will yet become.
Remember when Twitter used to give archives of tweets to the Library of Congress? And had a firehose for folks to consume as many tweets as they could?
My second reaction was “I wonder how hard it would be to take a list of twitter handles and scrape their feeds into an S3 bucket that’s fronted with a CDN.”
It's a shame that he's not able to escape his pathological belief that his product approach is the right approach, regardless of the grotesque impact he's making on what was a good thing for all of us.
I can't tell if it's different people talking now or if people really are that fickle. Probably a bit of both.
I hold both views.
The subtlety is in whether the User is beholden to Twitter Inc in a Just Along For The Ride sense, or is Creating a Burden for Innocent Others.
Former is most of us, latter is institutions incapable of acting insightfully considering the possible future (this one we now live in).
The more Musk makes changes, the faster it degrades.
Stuff like this will happen more often in the future. It's downhill all the time.
The last few months showed he’ll silence people whenever he wants (his right— it’s his platform— but clearly he lied)
And now you can’t even consume that limited speech on Twitter without letting Twitter track every single thing you read and link you click on.
There’s lots of legitimate reasons to criticize, but making a more efficient company and firing valueless employees is not one of them.
Seems just another day at Twitter dot com.
What should we do to stop that? I’m open to ideas."
https://twitter.com/elonmusk/status/1674898695534309378
"1. Scraping is already disallowed by T&C.
2. The scraping orgs dgaf & mask their IPs through proxy servers or through orgs that appear legit. For example, a recent massive scraping operation originating from Oracle IP addresses was just using their servers as a laundromat.
3. We absolutely will take legal action against those who stole our data & look forward seeing them in court, which is (optimistically) 2 to 3 years from now."
What does “our” refer to here? Does Twitter (i.e. musk) own the data in any sense? Or does he mean it as “we the people’s data”?
Very off-putting to read that sentence. Obviously he’s trying to monetize the user generated data in this LLM rush as other avenues to monetizations have flopped.
But I'm happy to speculate: Organizations violated the twitter TOS by scraping, and he's going to sue the organizations for it.
> In a second ruling in April 2022 the Ninth Circuit affirmed its decision.[5][6] In a November 2022 ruling the Ninth Circuit ruled that hiQ had breached LinkedIn's User Agreement and a settlement agreement was reached between the two parties. [7]
And he just as famously went crawling back to Twitter. He's too addicted to quit it.
he is hopelessly addicted to Twitter
That's probably close to 99% of HN users
Seriously though, lots of people don’t use Twitter. Of my friends, only one has a Twitter account, and he’s switched away since Musk took over.
In the beginning I felt the original length-limit fundamentally doomed it to a certain kind of not-so-valuable conversation. I mean, hell, even this quick comment here is already ~381 chars. I know the limits have been raised, but I think its effect on the culture remained.
Projection probably. I use my twitter account so seldom, I'm definitely in the 1%.
But then, I'm probably in the 0.5% who if HN happens to be down, will check back the following day instead of worrying.
There may be a post about it in /r/sysadmin
Would HN being down count as an emergency for anybody else?
https://syndication.twitter.com/srv/timeline-profile/screen-...
| |
"Nitter.cz is not working, just like all other Nitter instances. The reason is Twitter blocking all access to it's content without login.
We are sorry, but there is nothing we can do about it right now and we are not sure if the situation will change in the future.
Don't trust corporations, especially those where one egomaniac has all the power. Use open-source and community driven solutions if you can (like Mastodon).
Sincerely, NoLog.cz collective
PS: You can also donate to us to keep our other services running"
I'm never, EVER making a twitter account. However publishers still communicate with me via tweets I could see. Now that I need an account to view tweets, publishers just have a smaller audience.
I'll just see the screenshots on reddit anyway :)
If Musks claim today is true ("This platform hit another all-time high in user-seconds last week"), then it is very much not in decline...
[1] https://www.zdnet.com/article/twitter-seeing-record-user-eng...
Looking at the numbers that is a lot of debt:
https://www.wsj.com/articles/elon-musks-twitter-takeover-see...
"As part of the deal, Twitter will add about $13 billion of debt. Analysts estimate, based on terms previously laid out in documents related to the transaction, that Twitter would be on the hook for annual interest payments of more than $1 billion, compared with some $51 million in 2021."
Just because that's what was written into the contract when the acquisition happened. Something different could have been put in the contract, but this is what was actually written in it and signed off on by all the parties to the deal.
Twitter monetizes via 1. advertising, 2. subscriptions, 3. API sales, admittedly I have no idea of the actual numbers.
The first claim (that Twitter needs revenue urgently) seems false since the owner has deep pockets.
They became that rich by not losing money. It’s a sinking ship and he is trying to plug the hole.
Currently Twitter is in the red, and he needs it to not be that and to generate a multi-billion surplus to pay back the investors he took up to finance this.
Can you explain how Twitter manages to spend $20B per year?
Say they have 1000 employees @ $200k/year, that's $200m/year (1/5th of 1B). Where are the remaining $19.8B being spent? That figure doesn't pass the sniff test.
What I’m saying is for Elon to even break even on the deal, using the information we have access to, he needs its evaluations to reach at least $40B. Once it reaches that, it actually proves itself as having been at minimum a neutral purchase. Albeit the real numbers it has to reach and how much profit it needs are both currently unknown.
But more seriously, maybe the value is in it being a real-time source of new data, albeit the signal is drowned in noise like a needle in a hay stack.
This is truly a sad state of affairs.
For instance if you search "movies like X", on Google by default most results are automated aggregators or one critics perspective. If you search "movies like X reddit" you get more recommendations.
Y'all realize you can just not use reddit and Twitter right? You can host your own forum on the Internet
If it wasn’t so sad, this would actually be quite amusing to watch.
> Do you foresee more hosting providers offering one-click, fully managed, ActivityPub deployments?
Spez is inarguably incompetent as a leader and I realize that if I saw him in public, I would point and laugh.
Truly. I understand the fear of social media companies dealing with data scraping for rivals to run AI models, but it has been truly fumbled from a messaging perspective.
Is there a viable alternative for them to use now that twitter is going to block most of our county’s residents?
texts through registration on the site of the city hall, if cell braodcast isnt implemented by the phone company
public emergency sirens
radio
TV
actual website of the city hall
local newspaper
social fabric, with residents actually calling each others, for non emergency topics ?
(Hacker news ?)
Twitter has never been a great solution on its own. It could have contributed along with other channels of communication, so you should already have these.
One disadvantage with all these is they're one-way. Twitter had reasonable ways to manage one-to-many conversations that other solutions lack.
So of course it definitely could go badly fo a variety of reasons—but I think there's good reason to be optimistic that it will go well.
Mastodon has the mind share already as the de facto Twitter replacement.
Bluesky looks suspiciously like it’s dead in the water.
It did, until it got infested with far-left refugees who started banning entire instances whose users their have political disagreements with. Between that and the lack of ease of signing up or setting up your own instance, Mastodon is not taking over from Twitter anytime soon, although it has formed its own community that I'm sure will continue to use it. If Mastodon becomes even as popular as Tumblr now, let alone at its peak, I would consider that a surprising success.
As for the "tech community", some hackers did go over to Mastodon. That's great, but they didn't cause IRC or mailing lists to take over the world back in the day, and Mastodon is even more technically dysfunctional than those (making it so easy to completely block instances is a major issue). Call me when VCs start moving en masse to Mastodon.
I signed up, and poked around a bit.
But by and large I decided I was done with that public style of social media.
Now I’m pretty much just HN, and some Dark Social; Telegram, iMessage and FB Messenger friends groups.
Don’t @ me about FB Messenger LOL. I know…. But for our local School Comms it’s essential.
I live in an obscure part of Australia, and the less VC posts I see the better. If I saw a post from @jason I’d know it was time to leave ;-)
But seriously, for me, I have zero interest in any of those VC posts or commentary. The value is in more local individuals and other interest groups being active. YMMV :-)
The second they should already know about / be a part of. Maybe email them to check?
For the first, it really depends what area you're in. Australia/Victoria for example has a state-wide https://www.emergency.vic.gov.au/respond/ which has its own app with location/severity based alerting. There may be something similar around you.
And if it’s life or death, there’s Emergency Alert System. Goes on TV, radio, and straight to your phone, complete with the Cold War-era alert tone. I hope Twitter merely supplemented that?
Where I live, the tornado siren, which is beyond extremely loud and mounted high up on a pole in the highest point in the neighborhood, can be used to broadcast PSAs. It screams "THIS IS ONLY A TEST" at noon on the first Wednesday of every month. And you hear it whether you want to or not.
This actually happened to me, and guess where all of the official police updates were posted.
Youre not wrong, but that doesn't mean they make effective use of them. During one of Canada's deadliest mass shootings, the police elected to only post information on Twitter.
https://www.cbc.ca/news/canada/nova-scotia/rcmp-twitter-aler...
I'd think most of the power of posting emergencies on Twitter would be notifications?
My house was in the evacuation zone for most of the recent fires. I just kept this link open on my phone and refreshed it as needed.
Quite often they will tell you when to expect the next update, ie) within an hour
Even if I had a Twitter, I’d say the random notifications would be more frustrating than valueable in an emergency situation. If I’m in a safe enough spot to look for updates online I can pull up the local OEM Twitter.
https://syndication.twitter.com/srv/timeline-profile/screen-... add username at end
from https://einaregilsson.com/redirector/ and https://gist.githubusercontent.com/robotblake/a0f020381c1a91...
https://syndication.twitter.com/srv/timeline-profile/screen-...
change the word at the end with the username you want to see
(check out https://rsshub.app/twitter/user/elonmusk)
For the very desperate, this is still "working" for some reason: https://tweettunnel.com/ggreenwald
At the end of each snippet is a t.co link with a /{tweet id} at the end. You'll have to click on it and copy the status/tweet id from the url after the "%2Fstatus%" part. Paste that number into an embed link: https://platform.twitter.com/embed/Tweet.html?id=16763153424...
https://platform.twitter.com/embed/Tweet.html?id=16748657311...
I wonder if this is all a result of the API price hike. Folks probably tried to keep integrations alive by scraping, so now there is a lot of extra traffic on the main site. If integrations start acting like embedded calls, they will have the same problem there soon.
At the very least if these sites are being used for official communication that might be critical to peoples safety some sort of privileged status or ToS should be negotiated. Can you imagine Musk banning some random non-USA government agency because he had a fit while high at 3 am?
Also: https://xkcd.com/743/
Google results that require a login to view are cancerous
Not only, but also.
Does the Chinese government run an account? Have American politicians been as upset about this as they seem to have been about TikTok? I'd check the former, but, well, the subject under discussion.
Exactly this. It's fine if they use twitter to syndicate news that is also announced on official government systems, but not as a primary and certainly not as a solitary distribution method.
> Can you imagine Musk banning some random non-USA government agency because he had a fit while high at 3 am?
That would be hilarious and maybe it would result in some people learning that twitter is in fact a private corporation that can do whatever it wants, but i doubt it - similar incidents proved that large swathes of users believe twitter is or should be treated as public infrastructure rather than prompting significant moves to user controlled platforms
It’s a business. Free speech is the brand.
Reach.
IDK if anyone is using it as the sole method of communication.
But Twitter in practice has a much higher reach than every other method.
Stuff like "public transit line 17 out of service" being announced only on Twitter is completely par for the course.
REACH
Any talk about privacy awareness is invalidated when public sector entities endorse these platforms and encourage citizens to participate.
Any talk about the public sector not picking winners is a joke when they explicitly advertise and provide links on their websites to particular platforms.
We have normalized alot of abnormal stuff in the past decade...
It's the responsibility of the government to make information available where the citizens are. As the citizens moved from radio and TV to social media, the government followed.
Of course, running a website, which allows those quick edits requires with dealing with secure infrastructure, maybe apps dir field agents to write something etc. which they could all outsource to Twitter (and vendors of tools built around Twitter API for scheduled tweets etc.)
Its somewhat similar with public broadcasters (BBC, German ARD etc ) putting their content on YouTube (where consumers are) vs their sites (where they control it, including privacy concerns)
Simple, if you are not there you do not exist. Go and check how many of your contacts follow the oficial accounts of your local government.
Also, the common layman isn't anyone that suddenly gets the urge to check your official webpage, if you are lucky you are in their social media results.
People stopped hosting their own forums. Frankly, it's hard to not see why. The constant spam and people avoiding bans wasn't helpful - and modern forum software like Discourse is pure agony to set up and maintain if you don't know what you are doing. Not that forum software hasn't always been hard to set up, but the modern software stacks are particularly hard to manage. Also, what normal people see as good UX, in my experience, almost completely does not match what computer engineers and the average open-source contributor sees as good UX.
I'd like to know more about why Discourse is this way. Why the fark can't I just docker compose it up and running? I'm almost but not quite thinking about paying for a hosted Discourse solution since time is money. But why do they make it so hard?
Now I can never see myself going near the site again. Oh well.
So clearly calling musk a "pedo guy" is just an affectionate nickname, too.
1) discovery
If I can’t find a forum, it’s not much use to me!
2) single sign-on
I don’t want to make a million accounts for a million separate forums. I want one account that I can take with me to every forum.
Both of these issues were solved by Reddit. This was its major value-add. The user base brought all of the remaining value with them.
This is also the value of the fediverse, of which lemmy is a fine example. Yes, the technology is way more complicated than phpBB, but it solves these very challenging problems which present huge barriers to growing an individual forum (which must be overcome once again, every time for every forum).
This is akin to online dating. Nobody needed 5 dates a week 20 years ago and because it's an option now doesn't mean it was a problem than. Optionality is just a byproduct of social platforms.
If you're referring to the time when the "old forums" were new, that is a world which no longer exists: a world where Google search results were useful and not overloaded with paywalled sites and SEO spam. Today, you're going to have a very difficult time finding those "old forums" unless you know the exact name of what you're looking for. And forget about browsing.
As for "people today have too many options, back in the day we had fewer options and we were fine!" that's a very old argument you'll have a hard time convincing many people of.
Why do you want that? Why would you want your identity on a functional programming forum to be the same as on a Star Trek fans site and a furries meetup group?
The other alternative where you don't care if functional programming and Rust programming forums are on the same id is the issue without the option of a single account
Disagree. I'd frame it as the "popular web" as we know it is over. There's plenty of other space on the internet for other web experiences that are different from what corporations have given us over the past two decades.
I’ve been avoiding clicking Twitter links, and it’s frustrating to see them on the HN front page but then having to infer the content from comments.
Maybe HN could even have a “Submit a microblog link” feature for that kind of content?
When submitting a microblog link, there would be a small text box where you could copy the content of the tweet/whatever. This should be fair use as a quotation (but IANAL). It could be an important archival feature for situations like the current one, where all past HN submissions that point to Twitter are suddenly behind a login wall.
And I prefer the original source even if I don't like it.
---
I still have a few accounts I glance at from time to time. Hockey, Game Devs, Artists, etc. who haven't migrated away despite everything, so this is kinda obnoxious. I created some Redirector (https://einaregilsson.com/redirector/) rules to redirect Tweet and Twitter Profile URLs to their HTML embed equivalents.
Should be able to just import the rules and it seems to work alright with some caveats.
* I have no idea if this will continue to function.
* I've only tested some random links from my Discord and Slack groups.
* Profile links only show the most recent 20 tweets.
* Tweets will show quote-tweets, but no replies (though maybe that's a good thing).
* Obviously won't work for mobile.
Rules are at https://gist.github.com/robotblake/a0f020381c1a919cf9720f9ae...Edit: And just realized /u/justnotworthit posted the link earlier too.
I think Write Once, Publish Everywhere (including both centralised and federated) is much better.
twitter and facebook make automating that difficult with their API restrictions.
I find it tedious to update various social media platforms by hand, especially when each platform has its own rules and conventions. There are paid services that help but they often don't cover all of the platforms that I use, or are prohibitively expensive. Also if you just post a link to your site some social media platforms will treat you as a spammer.
Now I'm pretty sure your reach is severely limited if you don't pay; and now tweets are blocked behind a login.
How's it going, really?
" Roadmap and Growth
A few weeks ago, we intentionally slowed our invite roll-out while we built more moderation tooling and capacity for users on the app. We staffed a content moderation team with shifts that cover a 24/7 schedule, and consulted with trust and safety experts to establish new processes and policies to support a growing userbase. We’ve resumed sending out daily invites to the waitlist, which is where the majority of users already on Bluesky received their invites. For those who don’t know someone personally with an invite code, the waitlist is the fastest way to receive a code, though please be patient as we work through the list. "
blueskyweb.xyz
bsky.app
@username.bsky.social
My argument is:
Because they remaining main usage of twitter seem to be people "tweeting" small important (from their POV) bits of news. This could be political news but also e.g. announcements of content creators.
The changes in the algorithm already where not so grate for some people for this use case.
But most important for this use case is that this tweets need to be world readable even if not world discoverable.
With this gone a lot of people now have to look for a place where they can do world visible announcements, and at that point why not only use that place? Especially if it doesn't require you to e.g. buy something like twitter blue to increase the chance that people subscribed to you see your announcement.
I’ve wanted to do something similar to Mastodon where you can follow hashtags.
Sounds like you might have solved that?
(Who knows how long it will keep working, of course.)
I'm not sure why anyone wants "neutrality" to mean 50/50 airtime, instead of attempting to present the best available picture of the world, even if it's not favorable to one side.
It's also not a great list. Igor Sushko, for example, is very susceptible to posting unverified things that later turn out to be false.
A much better list is Noah Smith's at: https://twitter.com/i/lists/1492242776825552896. Yes, you now need to log in to see it, unfortunately.
Public Twitter lists seem like the sort of thing that's quite scrape-able though, even behind a login wall. With sufficient caching & semi-randomised access it should be fairly hard to detect which account is doing the scraping.
If you want the most up to date stuff you would probably have to follow some Russian and Ukranian channels, Telegram has built in Google translate to make that slightly easier. Reddit is probably a good place to start to find some.
While I'm interested in following the war, I don't have a desire to watch combat footage so I've been a little hesitant to jump into his telegram channel.
I kinda get the sense that it's probably not super easy to find a telegram channel that doesn't post war footage.
You don't miss out on much though. Twitter just seems to be random hot takes and celebrity nonsense.
Basically this vision is one where everyone run's Google Reader in their browser and, if you're really gung-ho, you republish via some VPS hosted thing. The structure is a little like DNS - a hierarchical database.
Or you try to ride the coattails of network effects and remove any obstacles to your information, index public data, and generally tries to weave in the service in the public web as tight as possible. This is the www model, as exemplified by Google.
Then there are the ones in between, such as newspapers, who generally wants to hide their information behind their service to increase the perceived value, but also wants to be reachable from the public web because without findability they have no value. This is of course paradoxical reasoning with no end of problems and no successful businesses.
Twitter from the start was part of the web 2.0 movement with a clear www strategy. Now they are doing a strategic shift. But how valuable is twitter when they're no longer indexed by services such as Google?
I can not see how they can grow their business in the long term from this. They lose the sentiment of a mirror of public discourse and instead will become like any web forum. So there will be zero journalistic value, and why would celebrities want to hang around then?
It's not even a slippery slope. It's just a slope.
The most recent updates are available on that page, but for anything older than that you'll need a Twitter account to read stripestatus on Twitter. Even their Atom feed is a useless river of "A status update was posted" titles with truncated tweet content and t.co links.
It's such a shame, Twitter used to be so useful for stuff like this.
import datetime
import json
import os
import sys
out = sys.stdout
media = {}
for med in os.listdir("data/tweet_media"):
tid = med.split("-")[0]
exist = media.get(tid, [])
exist.append(med)
media[tid] = exist
out.write("<html><body>\n")
with open("data/tweet.js") as d:
vals = json.load(d)
for val in sorted(
vals,
key=lambda val: datetime.datetime.strptime(
val["tweet"]["created_at"], "%a %b %d %H:%M:%S %z %Y"
),
):
tweet = val["tweet"]
out.write("<div>\n")
out.write(f"<i>{tweet['created_at']}</i>\n")
for fname in media.get(tweet["id"], []):
if fname.endswith("mp4"):
out.write(
f'<video controls><source src="data/tweet_media/{fname}" type="video/mp4"></video>\n'
)
else:
out.write(f'<img src="data/tweet_media/{fname}"/>\n')
if "full_text" in tweet:
out.write(f'<p>{tweet["full_text"]}</p>\n')
out.write("</div>\n")
out.write("</body></html>\n")Now Twitter.
I am better off for these changes.
You can be too. Just log out. It will pass.
Apparently, any kind of access is now deemed adversarial. ;-)
If the fediverse ever had this kind of scraping pressure, it would collapse in minutes.
edit: Scrolling a thread counts as a hit against this every time it needs to load more tweets, so this is very easy to hit.
(And Twitter won't want to allow caches or archives since if course they could be used to get around read rate limits)
To the extent that tweets (does Twitter not call them that anymore?) are an important part of the public record, including being able to verify whether someone really did tweet a certain thing at a certain time, this is a little bit scary.
And just generally for our ability to preserve historical record, already threatened by the digital world, and apparently getting worse instead of better.
Is there an acceptable threshold of free viewing before it becomes abusive? (Think, getting a single free See's candy from the store vs. employing an army of people to source thousands of pounds of chocolate treats.)
With the Reddit API issue I'm honestly unsure where I stand. I love(d) Apollo and want it to succeed, but Reddit is doing the work and not getting the rewards. Where do you draw the line at "fair"?
In fact I think this is good. It makes it very clear that no, it’s not your content and no, you don’t deserve any rights just because you feel like you own it. Twitter will do as it pleases with “your” content.
I would very much welcome a far more informed environment where people were forced to face the details of IP rights and what it means to post content on these services.
1. Is it worthwhile giving free labour to these services by generating their content,
and,
2. Do we want to login and/or pay to check user generated content?
Personally, I think the answer will be no. If these services generated their own content, which we actually value, then it would be a different matter. But they don’t, so if they put up road blocks I suspect they will get a bit of a shock to learn they aren’t actually as vital to society as they thought they were.
Surely now is the time to capitalise on the discontent on Twitter, even if it means dealing with a few scaling pains.
I don't know for sure if it is possible to fight it. Centralized platforms will always win because they are very convenient and UX is very smooth while user plays ball by platform's rules. Migrating to decentralized is mostly for geeks.
I am geek myself, but I still can't find energy to migrate at leas email to my own domain. Instead I keep using GMail, despite all the risks (I am citizen of Russia, so in my case risk of deplatforming from Google is even higher).
They put search behind login and now tweets themselves to try to push people to use their API. Twitter is still the bigest source used in public sentiment analysis and it's still relatively easy to scrape with login so this just shifts everything to grey/dark markets. The essentially means public researchers, students w/e lost access while dark market value increases significantly.
So, as per character, this just empowers corporations and punishes public. Hopefully this means the upcoming end of Twitter.
(just for the statistics, I know it does not matter what is problem for me specifically and what not, just one point of view added. a person like me nervous about tracking my internet activity and blocks embedded twitter, does not like the twitter lifestyle of brief, single viewpoint, and oppinionated texts so never goes there by will, there will be no change. well written articles are much better anyway than brief and sudden reactions the twitter is made for, basicly smalltalk before the real conversation.)
The final act will be when they kill Twitter embeds on the basis of too many people are now viewing Tweets but not signing up for Twitter.
Literally doing the latter thing - killing embeds - would've been a more sensible decision.
They have become an anti-thesis of freedom that was promised when World-wide-web emerged.
She is much more akin to a CEO at Large running around closing ad deals.
This has been Twitter's Modus operandi from well before Musk acquired them.
They announced new features through Tweets (remember Fleets?) with glaring issues at launch (no authentication required to view a Fleet) and without any corresponding API documentation. The GraphQL that backs their website has more features and bug fixes than their paid API!
They also famously do not have a staging environment, which frequently caused issues when new features were deployed or tested with a subset of users.
Twitter has never been a trustworthy or reliable partner. They are just more of an obvious dumpster fire since the Musk acquisition.
Twitter's got its share of issues, but not having staging isn't one of them because there's not much value for massive distributed systems like that to have a staging environment, not relative to the effort required, anyway.
I wonder what happened to embedded tweets?
Definitely disturbing that journalists (especially) figured it was good archival practice to rely on the Twitter API in providing context.
For years, if I didn't enable twitter's javascript, news articles are missing images and quotes, obviously so. It's embarrassing, I honestly don't know how they recover from this, I don't know why they kept relying on Twitter embedding when screenshots and copy/paste work better and don't break.
Very innovation, many progress.
Here's the example link from that comment: https://platform.twitter.com/embed/Tweet.html?id=16748657311...
Edit: and here's a random news article (post?) that has a working embedded tweet: https://www.theverge.com/2023/6/30/23780357/new-footage-of-t...
I just made an updated tool that should work Replace http://twitter.com by http://traittor.net
https://twitter.com/elonmusk/status/1675187969420828672
to
https://traittor.net/elonmusk/status/1675187969420828672
Is normally bypassing no registered account and limit by day.
Have fun :)
For my part, this is an annoyance since I use nitter's API to feed Tweets to a Slack I share with friends.
Nearly everybody is going to be getting off of Twitter now. It no longer is what it used to be. A place where celebrities could brag about themselves publicly.
Remember, this is what influenced Facebook to open up and make all of the profiles public.
Mastodon and blue sky here we come.
There will probably also be tons more abandoned accounts halfway through the registration process, as the last time I remember trying to register one, I was asked to provide a phone number and noped out immediately.
Elon has crashed a lot of my faith in him as anything more than another insane industrialist. Fuck Elon, fuck Twitter
I‘m not really sad about that change. Just going to miss out some things because I don‘t see why I should register to read a few tweets a week.
Does Internet Archive capture twitter content? If so how will all of these changes impact them?
This seems very unlikely. If they really just wanted to stop just AI tools from searching twitter, it would be very easy to prevent them from doing it at scale by imposing basic rate limiting and device intelligence (or even something like the puzzle LinkedIn makes you solve before viewing someone's profile while not logged in).
I'm very confused as to why they may not want unlogged-in human lurkers who are still seeing and clicking on ads when on the Twitter website.
Same issue with Reddit, it's a false excuse for embarking on some other kind of cash-grab policy.
They claimed the almost-no-warning API changes were necessary to stop the "AI", except that all the big (and therefore significant) actors could have been stopped by a change to the terms of service or some modest rate-limits.
Seems like it would be but didn't know if it works using scraping or something.
Musk and his team are zigzagging to evaluate possible business models that make money.
having normal user account throttled in desktop browser the whole day today so far
great job twitter !
This anti pattern exists all over.
Both results are positive IMO
Good to know I'll never visit the dumpster fire of Twitter again.
we've come full circle
https://twitter.com/elonmusk/status/1674887204580073474?s=20 ---- @elonmusk
"Several hundred organizations (maybe more) were scraping Twitter data extremely aggressively, to the point where it was affecting the real user experience.
What should we do to stop that? I’m open to ideas. 5:06 PM · Jun 30, 2023"
https://twitter.com/elonmusk/status/1674942336583757825?s=20 ---- @nearcyan
"so @elonmusk now that twitter blocks all requests that are not logged in, tweets can no longer be embedded in most chat apps
I'd strongly suggest reconsidering the UX+growth tradeoffs made here (take note how tiktok, youtube, etc, do not need this despite much higher b/w req!)"
----
@elonmusk
"This will be unlocked shortly. Per my earlier post, drastic & immediate action was necessary due to EXTREME levels of data scraping.
Almost every company doing AI, from startups to some of the biggest corporations on Earth, was scraping vast amounts of data.
It is rather galling to have to bring large numbers of servers online on an emergency basis just to facilitate some AI startup’s outrageous valuation."
- They are hard to use for regular people due to lack of modern and intuitive clients.
- They are somewhat hard to discover and learn. For example, browsers don't indicate the presence of RSS feeds without plugins anymore.
- Security and encryption are an after-thought, if any
Almost all of the above are due to the stagnation of their development (protocols and clients) due to emergence of easy-to-use, but centralized alternatives. They are not good modern options as they are now. But they are good models to build more modern alternatives on.
Several hundred organizations (maybe more) were scraping Twitter data extremely aggressively, to the point where it was affecting the real user experience. What should we do to stop that? I’m open to ideas.
Much like paywalled sites, can we please stop posting twit links now?
There was a point when Twitter was good enough that maybe they could have pulled something like this and gotten away with it. At this point, I think all this will do is hasten their irrelevancy.
Even the content aside (that you have to wade through), just from a technical perspective the Twitter experience leaves a lot to be desired.
Putting the internet in the hands of the Government wouldn't fair much better.
The Internet is supposed to be distributed. We've gotten so used to consolidated services that we have forgotten this lesson.
Sounds like the issues we currently have with democracy.
Centralizing control of something that was designed to be distributed
This is human nature/greed unfortunately. Look at any natural (distributed) resource. The current economic system rewards this as well.It’s also not a very interesting problem to solve because of the type of cliffs you will run into due to precisely how the “internet works”
What is relevant is governance. We allow billionaires and venture capitalists to govern a commons that we all rely on. Surprise surprise, it isn't going well.
The solution is not to have (difficult to scale) federated alternatives. The solution is collective ownership.
Imagine for a moment that the multinationals that are increasingly in charge of our lives were owned by their customers. Imagine they had a fair electoral system, reflecting the variety of those users, limiting them to one person, one vote, and that their constitutions were designed to guarantee the rights of minorities.
The journey that most countries went on through the 20th and 21st centuries, in other words.
Tech giants and other multinationals are a different kind of beast, because they govern a little slice of our lives instead of having carte blanche. But it is not beyond the realm of possibility for democratically operated multinationals to exist. It will be hard to do, but IMO, that approach has a bright future because non-techies can grasp it and participate in it more easily, and that is one less barrier to a runaway network effect than the fediverse has.
A dynamic, user-driven community still thrives in the vast expanse of the digital world, yet it lies hidden beyond the towering edifices of corporate-controlled structures. Discovering these spaces has become an increasingly formidable task, as the infusion of corporate social content into journalistic and blogging platforms perpetuates the mirage that such networks are all that exist.
Each colossal tech corporation we see today began its journey as a modest, affable endeavor. As these projects expanded with their burgeoning popularity, users neglected to challenge the escalating influence and control these companies wielded.
Nitter was merely an alternative facade to Twitter. Despite offering an ad-free environment, it lacked substantial advantages as the underlying platform remained the same - Twitter.
However, the digital realm is not void of choices. Federated social media is emerging as a profound alternative. Yet, a majority of those voicing concerns about corporate social media seem to dismiss options like Mastodon. This is primarily due to their increased technological demands and people's comfort in having a corporation guide their online journey.
The power to reshape your digital footprint rests in your hands. You can sever ties with your corporate social media accounts. You can choose to eschew media that incessantly embeds corporate social media content. You can advocate for an internet not ruled by corporate influence. All it requires is the willingness to venture beyond the realm of comfort.
I have a hard time reconciling this perspective with history. Were any of these ideals present among the people/organizations responsible for the internet and the Web at the time that they were being developed? Or is sentiment like yours something that people adopted later on?
The internet was conceived as a DARPA concept of reliable government communications in the face of unreliable transport, among many other research interests. For most of its early existence (through at least the NSFnet incarnation in the US), it was the private preserve of academic, government and military users, along with some of the corporations that supported them and commercial use beyond supporting projects was prohibited (e.g. you couldn't use it for advertising). It was far from a 'Democratic haven'. There were epic flamewars over 'do we let any more commercial content in our private backyard?' and 'why would we let regular people in?'.
However, a pervasive dip in technological literacy...
Right...it's bad we let the proles into utopia. So much for 'democratic havens'.
If you can't get the basic background right, it really damages the credibility of the rest of the screed (which I mostly agree with).
FOSS, fediverse, IPFS all had their chance, and they blew it. Corporations were the ones who opened up the internet to the 99% of people who would otherwise never have been there at all, and now they want to collect their cut.
Even now, these federated sites on the rise have technicial growing pains. And those will probably take years to get through until it's to a point where everyone can use it with little friction.
Twitter made a convenient, easy to use, centralized (which is an absolute positive for user experience), social media product that attracted people, by their own free will. The number of people using a social media service amplifies its "usefulness", so the more people, the stronger it attracts new users.
We didn't put the internet in the hands of these corporations. We walked over and sat in their, easy to use, hands.
I don't think they'd have ever bothered inventing privacy violating trackers A/B testing (though if I'm wrong this is the best place to assert wildly and be quickly corrected).
Without tech companies, it was still immensely useful.
Apps on iPhones? Great, but the internet doesn't need them to be hugely culturally and socially important — and I'm saying that as an iPhone app developer since before the first iPad came out.
We the people were inactive & didn't figure out how to weave together our individual & community sites to create a compelling multi-party space.
Or we could try to create alternative centralized but non-corporate systems. Not sure what other options there are.
I don't like where we are either. But new power has to be created. Hard work of figuring out protocols to converse across & usefully home our content/words on is sort of just beginning.
Federated systems are a nice idea, but they're not funded and will crumble under the same pressure until they too go into private mode. It's simply not a financially sound decision to run an open node that is continually harvested by corporations seeking to profit off the conversations occurring on your platforms.
I think it's a mistake to block third party providers from profiting from their service. First of all the hypocrisy in that all of these large companies exist because of massively profiting off of mostly uncompensated user-created content. Second, this drive toward relentlessly monetizing every aspect of your company's business is how we get degraded services like Microsoft putting ads in their search bar. It's one thing if a downstream OEM does it, I can just use an alternative OEM that doesn't shovelware the crap out of their product. But when the primary provider does it, then their service is permanently borked and eventually becomes unusable. So if the idea of blocking scrapers is because the goal is to eventually provide their own shitty AI services, I think Twitter et al are just going to end up killing their own geese -- making their own services unusable out of greed.