Now with LLMs, I can't remember the last time I visited StackOverflow.
The harder the problem, the less engagement it gets. People who spend hours working on your issue are rewarded with a single upvote. Meanwhile, "how do I concat a string" gets dozens or hundreds of upvotes.
The incentive/reward structure punished experienced folks with challenging/novel questions.
Pair that with the toxic moderation and trigger-happy close-votes, you get a zombie community with little new useful content.
Eventually SO becomes a site exclusively for lurkers instead of a platform for active participation
This is literally not true. The rate you learn and encounter new things depends on many things: you, your mood, your energy etc. But not on the amount of your experience.
> The harder the problem, the less engagement it gets. People who spend hours working on your issue are rewarded with a single upvote.
This is true, but not relevant, I don't think many people care. Some might, but not many.
The questions you land on will be unanswered or have equally confused replies; or you might be the one who's asking a question instead.
I've "paid back" by leaving a high quality response on unanswered SO questions that I've had to figure out myself, but it felt quite thankless since even the original poster would disappear, and anyone who found my answer from Google wouldn't be able to give me an upvote either.
When someone says "I feel like" and you answer "No, you don't", you're most certainly wrong :-).
I do feel like the parent.
Are you being serious here?
I was used to doing that, but then the moderation got in the way. So I stopped.
I've answered about 200 questions. I've asked two, and both remain unanswered to this day. One of them had comments from someone who clearly was out of their league but wanted to be helpful. The people who could've answered those questions are not (or were not at that time) on SO.
If the moderators are not competent to understand if your question is a duplicate or not, and close it as duplicate when in doubt, then it contributes to the toxic atmosphere, maybe?
They could go with "when in doubt, keep the duplicate", but they chose the opposite. Meaning that instead of happy users and duplicates, they have no duplicates, and no more users.
My initial (most popular) questions (and I asked almost twice as many questions, as I gave answers) were pretty basic, but they started getting a lot more difficult, as time went on, and they became unanswered, almost always (I often ended up answering my own question, after I figured it out on my own).
I was pretty pissed at this, because the things I encountered, were the types of things that people who ship, encounter; not academic exercises.
Tells me that, for all the bluster, a lot of folks on there, don't ship.
LLMs may sometimes give pretty sloppy answers, but they are almost always ship-relevant.
"Sunsetting Jobs & Developer Story" 3/2022 https://meta.stackoverflow.com/questions/415293/sunsetting-j...
(That seems comparable to arguing that Facebook shouldn't subsidize posting baby photos).
But if it was the case that SO mgmt decided (2017-2020) that they didn't care to keep experienced users engaged, and just let the site degenerate into new users posting bigger volumes of duplicates, questions without code, etc., then that would be on them. You don't have to assume their actions were rational; look how badly they mismanaged moderation in that period and how many experienced users that lost them.
I don't use LLMs eother. But the next generation might feel differently and those trends mean there's no new users coming in.
Which is kinda cool, but also very biased for older contributors. I could drop thousands of points bounty without thinking about it, but new users couldn't afford the attention they needed.
This is killer feature of LLMs - you will not became more experienced.
>zombie community
Like Reddit post 2015.
For programming my main problem with Reddit is that the quality of posts is very low compared to SO. It's not quite comparable because the more subjective questions are not allowed on SO, but there's a lot of advice on Reddit that I would consider harmful (often in the direction of adding many more libraries than most people should).
Gen 1: stackoverflow.com (2008)
Gen 2: chatgpt.com (2022, sort of)
Google answers
And the horrific Quora
Random example:
http://answers.google.com/answers/threadview/id/762357.html
It's remarkable how similar in style the answers are to what we all know from e.g. chatgpt.
No way.
Proof: https://web.archive.org/web/19990429180417/http://www.expert...
I think overall SO took the gamification, and the “internet points” idea, way too far. As a professional, I don’t care about Reddit Karma or the SO score or my HN karma. I just wanted answers that are correct, and a place to discuss anything that’s actually interesting.
I did value SO once as part of the tedious process of attempting to get some technical problem solved, as it was the best option we had, but I definitely haven’t been there since 2023. RIP.
I disagree, I always thought it SO did a great job with it. The only part I would have done differently would be to cap the earnable points per answer. @rndusr124 shouldn't have moderation powers just because his one and only 2009 answer got 3589 upvotes.
You ask how to do X.
Member M asks why you want to do X.
Because you want to do Y.
Well!? why do you want to do Y??
Because Y is on T and you can't do K so you need a Z
Well! Well! Why do you even use Z?? Clearly J is the way it is now recommended!
Because Z doesn't work on a FIPS environment.
...
Can you help me?
...
I just spent 15 minutes explaining X, Y and Z. Do you have any help?
...(crickets)
I see it all the time professionally too. People ask "how do I do X" and I tell them. Then later on I find out that the reason they're asking is because they went down a whole rabbit hole they didn't need to go down.
An analogy I like is imagine you're organising a hike up a mountain. There's a gondola that takes you to the top on the other side, but you arrange hikes for people that like hiking. You get a group of tourists and they're all ready to hike. Then before you set off you ask the question "so, what brings you hiking today" and someone from the group says "I want to get to the top of the mountain and see the sights, I hate hiking but it is what it is". And then you say "if you take a 15 minute drive through the mountain there's a gondola on the other side". And the person thanks you and goes on their way because they didn't know there was a gondola. They just assumed hiking was the only way up. You would have been happy hiking them up the mountain but by asking the question you realised that they didn't know there was an easier way up.
It just goes back to first principles.
The truth is sometimes people decide what the solution looks like and then ask for help implementing that solution. But the solution they chose was often the wrong solution to begin with.
I spent years on IRC, first getting help and later helping others. I found out myself it was very useful to ask such questions when someone I didn't know asked a somewhat unusual question.
The key is that if you're going to probe for Y, you usually need to be fairly experienced yourself so you can detect the edge cases, where the other person has a good reason.
One approach I usually ended up going for when it appeared the other person wasn't a complete newbie was to first explain that I think they're trying to solve the wrong problem or otherwise going against the flow, and that there's probably some other approach that's much better.
Then I'd follow up with something like "but if you really want to proceed down this rrack, this is how I'd go about it", along with my suggestion.
I don't think your analogy really helps here, it's not a question. If the question was "How do I get to the top of the mountain" or "How do I want to get to the top of the mountain without hiking" the answer to both would be "Gondola".
Except that SO has a crystal clear policy that the answer to questions should be helpful for everybody reaching it through search, not only the person asking it. And that questions should never be asked twice.
So if by chance, after all this dance the person asking the question actually needs the answer to a different question, you'll just answer it with some completely unrelated information and that will the the mandatory correct answer for everybody that has the original problem for any reason.
Yep. The magic question is "what are you trying to accomplish?". Oftentimes people lacking experience think they know the best way to get the results they're after and aren't aware of the more efficient ways someone with more experience might go about solving their problem.
...
Well, the pump at the gas station doesn't fit in my car, but they sold me a can with a spout that fits in my car.
...
It's tedious to fill the can a dozen times when I just want to fill up my gas tank. Can you help me or not?
...
I understand, but I already bought the can. I don't need the "perfect" way to fill a gas tank, I just want to go home.
> Is there any way to force install a pip python package ignoring all its dependencies that cannot be satisfied?
> (I don't care how "wrong" it is to do so, I just need to do it, any logic and reasoning aside...)
https://stackoverflow.com/questions/12759761/pip-force-insta...
Imagine a non-toxic Stack Overflow replacement that operated as an LLM + Wiki (CC-licensed) with a community to curate it. That seems like the sublime optimal solution that combines both AI and expertise. Use LLMs to get public-facing answers, and the community can fix things up.
No over-moderation for "duplicates" or other SO heavy-handed moderation memes.
Someone could ask a question, an LLM could take a first stab at an answer. The author could correct it or ask further questions, and then the community could fill in when it goes off the rails or can't answer.
You would be able to see which questions were too long-tail or difficult for the AI to answer, and humans could jump in to patch things up. This could be gamified with points.
This would serve as fantastic LLM training material for local LLMs. The authors of the site could put in a clause saying that "training is allowed as long as you publish your weights + model".
Someone please build this.
Edit: Removed "LLMs did not kill Stack Overflow." first sentence as suggested. Perhaps that wasn't entirely accurate, and the rest of the argument stands better on its own legs.
"Troubleshooting / Debugging" is meant for the traditional questions, "Tooling recommendation", "Best practices", and "General advice / Other" are meant for the soft sort of questions.
I have no clue what the engagement is on these sort of categories, though. It feels like a fix for a problem that started years ago, and by this point, I don't really know if there's much hope in bringing back the community they've worked so hard to scare away. It's pretty telling just how much the people that are left hate this new feature.
[1] https://meta.stackoverflow.com/questions/435293/opinion-base...
- A huge number of developers will want to use such a tool. Many of them are already using AI in a "single player" experience mode.
- 80% of the answers will be correct when one-shot for questions of moderate difficulty.
- The long tail of "corrector" / "wiki gardening" / pedantic types fill fix the errors. Especially if you gamify it.
Just because someone doesn't like AI doesn't mean the majority share the same opinion. AI products are the fastest growing products in history. ChatGPT has over a billion MAUs. It's effectively won over all of humanity.
I'm not some vibe coder. I've been programming since the 90's, including on extremely critical multi-billion dollar daily transaction volume infra, yet I absolutely love AI. The models have lots of flaws and shortcomings, but they're incredibly useful and growing in capability and scope -- I'll stand up and serve as your counter example.
It's very tedious as the kind of mistakes LLMs make can be rather subtle and AI can generate a lot of text very fast. It's a sisyphean taks, I doubt enough people would do it.
The whole pitch here just feels like putting gold flakes on your pizza: expensive and would not be missed if it wasn't there.
Just to say, I'm maybe not as experienced and wise I guess but this definitely sounds terrible to me. But whatever floats your boat I guess!
You oversimplified and lost too much precision. Try again?
I don't know why you put "duplicates" in quotation marks. Closing a duplicate question is doing the OP (and future searchers) a service, by directly associating the question with an existing answer.
Isn't this how Quora is supposed to operate?
Mind you, while I'm a relative nobody in terms of open source, I've written everything from emulators and game engines in C++ to enterprise apps in PHP, Java, Ruby, etc.
The consistent issues I've encountered are holes in documentation, specifically related to undocumented behavior, and in the few cases I've asked about this on SO, I received either no response and downvotes, or negative responses dismissing my questions and downvotes. Early on I thought it was me. What I found out was that it wasn't. Due to the toxic responses, I wasn't about to contribute back, so I just stopped contributing, and only clicked on an SO result if it popped up on Google, and hit the back button if folks were super negative and didn't answer the question.
Later on, most of my answers actually have come from Github,and 95% of the time, my issues were legitimate ones that would've been mentioned if a decent number of folks used the framework, library, or language in question.
I think the tl;dr of this is this: If you can't provide a positive contribution on ANY social media platform like Stack Overflow, Reddit, Github, etc. Don't speak. Don't vote. Ignore the question. If you happen to know, help out! Contribute! Write documentation! I've done so on more than one occasion (I even built a website around it and made money in the process due to ignorance elsewhere, until I shut it down due to nearly dying), and in every instance I did so, folks were thankful, and it made me thankful that I was able to help them. (the money wasn't a factor in the website I built, I just wanted to help folks that got stuck in the documentation hole I mentioned)
EDIT: because I know a bunch of you folks read Ars Technica and certain other sites. I'll help you out: If you find yourself saying that you are being "pedantic", you are the problem, not the solution. Nitpicking doesn't solve problems, it just dilutes the problem and makes it bigger. If you can't help, think 3 times and also again don't say anything if your advice isn't helpful.
Joel promised the answering community he wouldn't sell SO out from under them, but then he did.
And so the toxicity at the top trickled down into the community.
Those with integrity left the community and only toxic, selfcentered people remained to destroy what was left in effort to salvage what little there was left for themselves.
Mods didn't dupe questions to help the community. They did it to keep their own answers at the top on the rankings.
"Knowledge should be free" they said. "You shouldn't make money off stuff like this," they said.
Plenty of links and backstory in my other comments.
The timeline also matches:
https://github.blog/changelog/2020-12-08-github-discussions-...
https://github.blog/news-insights/product-news/github-discus...
It would have been super trivial to fix but google didn’t.
My pet theory was that google were getting doubleclick revenue from the scrapers so had incentives to let them scrape and to promote them in search results.
Reminds me of my most black-hat project — a Wikipedia proxy with 2 Adsense ads injected into the page. It made me like $20-25 a month for a year or so but sadly (nah, perfectly fairly) Google got wise to it.
They are closed for good reasons. People just have their own ideas about what the reasons should be. Those reasons make sense according to others' ideas about what they'd like Stack Overflow to be, but they are completely wrong for the site's actual goals and purposes. The close reasons are well documented (https://meta.stackoverflow.com/questions/417476) and well considered, having been exhaustively discussed over many years.
> or being labeled a duplicate even though they often weren’t
I have seen so many people complain about this. It is vanishingly rare that I actually agree with them. In the large majority of cases it is comically obvious to me that the closure was correct. For example, there have been many complaints in the Python tag that were on the level of "why did you close my question as a duplicate of how to do X with a list? I clearly asked how to do it with a tuple!" (for values of X where you do it the same way.)
> a generally toxic and condescending culture amongst the top answerers.
On the contrary, the top answerers are the ones who will be happy to copy and paste answers to your question and ignore site policy, to the constant vexation of curators like myself trying to keep the site clean and useful (as a searchable resource) for everyone.
> For all their flaws, LLMs are so much better.
I actually completely agree that people who prefer to ask LLMs should ask LLMs. The experience of directly asking (an LLM) and getting personalized help is explicitly the exact thing that Stack Overflow was created to get away from (i.e., the traditional discussion forum experience, where experts eventually get tired of seeing the same common issues all the time and all the same failures to describe a problem clearly, and where third parties struggle to find a useful answer in the middle of along discussion).
Often, doing what your users want leads to success. Stamping authority over your users, and giving out a constant air of "we know better than all of you", drives them away. And when it's continually emphasized publicly (rather than just inside a marketing department) that the "mission" and the "policy" are infinitely more important than what your users are asking for, that's a pretty quick route to failure.
When you're completely embedded in a culture, you don't have the ability to see it through the eyes of the majority on the outside. I would suggest that some of your replies here - trying to deny the toxicity and condescension - are clearly showing this.
You misunderstand.
People with accounts on Stack Overflow are not "our users".
Stack Exchange, Inc. does not pay the moderators, nor high-rep community members (who do the bulk of the work, since it is simply far too much for a handful of moderators) a dime to do any of this.
Building that resource was never going to keep the lights on with good will and free user accounts (hence "Stack Overflow for Teams" and of course all the ads). Even the company is against us, because the new owners paid a lot of money for this. That doesn't change what we want to accomplish, or why.
> When you're completely embedded in a culture, you don't have the ability to see it through the eyes of the majority on the outside.
I am not "embedded in" the culture. I simply understand it and have put a lot of time into its project. I hear the complaints constantly. I just don't care. Because you are trying to say that I shouldn't help make the thing I want to see made.
> trying to deny the toxicity and condescension
I consider the term "toxicity" more or less meaningless in general, and especially in this context.
As for "condescension", who are you to tell me what I should seek to accomplish?
This is a great example of a question that should not be closed as a duplicate. Lists are not tuples in Python, regardless of how similar potential answers may be.
If you imagine that the answer should be re-written from scratch to explain that the approach will be the same, you have fundamentally misunderstood the purpose of the site. Abstraction of contextually unimportant details is supposed to be an essential skill for programmers.
My favorite feature of LLMs, is the only dumb question, is the one I don't ask.
I guess someone could train an LLM to be spiteful and nasty, but that would only be for entertainment.
Hacker News, and we who frequent it, ought to have that in mind.
It's not like only slimy people get to use moderator tools like on Reddit, since you need a lot of reputation points you get by having questions and answers voted up. It's more like (1) you select people who write surface-level-good answers since that's what's upvoted, and they moderate with a similar attitude and (2) once you have access to moderator tools you're forced to conform with (1) or your access is revoked, and (3) the company is completely incompetent and doesn't give a shit about any of this.
We’re talking about how communities can become toxic. How we humans sometimes create an environment that is at odds with our intentions. Or at least what we outwardly claim to be our intentions.
I think it is a bit sad when people feel they have to be compensated to not let a community deteriorate.
The answer to all of these questions is yes, for the most part. Volunteers are much harder to wrangle than employees and it's much easier for drama and disagreements to flare when there are zero consequences other than losing an unpaid position, particularly if anonymity is in the mix.
Volunteers can be great but on average they're going to be far harder to manage and far more fickle than employees.
What makes a community worthwhile is its ability to resolve differences productively. I think that if you replace individual responsibility with transactionality you have neither community nor long term viability or scalability.
Then again, we live in times when transactional thinking seems to dominate discourse.
I didn't say it's not worth doing but it will bring challenges that wouldn't exist with employees. Paying people adds a strong motivator to keep toxic behaviour at bay.
Your experiences will heavily depend on the type of project you're running but regardless, you can't hold volunteers, especially online, to the same expectations or standards as employees. The amount of time and effort they can invest will wax and wane and there's nothing you can do about it. Anonymity and lack of repercussions will eventually lead to drama or power struggles when a volunteer steps out of line in a way that they wouldn't in paid employment. There is no fix that'll stop occasional turbulence, it's just the way it is. Not all of your volunteers will be there for the greater good of your community.
Again, that is absolutely not to say that it can't be worth the effort but if you go into it eyes open, you'll have a much better time and be able to do a better job at heading off problems.
I've seen other people express similar opinions to yours and it wasn't until they experienced being in the driver's seat that they understood how difficult it is.
Really, if we could apply some RLHF to the Stack Overflow community, it would be doing a lot better.
But LLMs get their answers from StackOverflow and similar places being used as the source material. As those start getting outdated because of lack of activity, LLMs won't have the source material to answer questions properly.
they’re pretty good at getting info that is very up to date by using tools to access the web
Yeah that's a charitable way to phrase "perform distributed denial of service attacks". Browsing github as a human with their draconian rate limits that came about as a result of AI bots is fucking great.I don't run personal sites worth millions of dollars. I do, however, use sites like Sourcehut, DigiKey, Github, Mouser, Farnell, etc, etc, etc. that have opted to put everything behind bullshit captchas because of the DDoS (nee AI) bots.
If everything would be well documentated SO wouldn't have being as big as it was in the first place.
LLMs were not productified in a meaningful way before ChatGPT in 2022 (companies had sufficiently strong LLMs, but RLHF didn't exist to make them "PR-safe"). Then we basically just had to wait for LLM companies to copy Perplexity and add search engines everywhere (RAG already existed, but I guess it was not realistic to RAG the whole internet), and they became useful enough to replace StackOverflow.
Their tendency to bullshit is still an issue, but if one maintains a healthy skepticism and uses a bit of logic it can be managed. The problematic uses are where they are used without any real supervision.
Enabling human learning is a natural strength for LLMs and works fine since learning tends to be multifaceted and the information received tends to be put to a test as a part of the process.
(Not to mention the confabulation. Making up API method names is natural when your model of the world is that the method names you've seen are examples and you have no reason to consider them an exhaustive listing.)
The worst thing with Q/A sites isn't they don't work. It's that they there are no alternatives to stackoverflow. Some of the most upvoted answers on stackoverflow prove that it can work well in many cases, but too bad most other times it doesn't.
Quicker than searching the entirety of Google results and none of the attitude.
For now. They still need to be enshitted.
So we’ll end up with a choice of low-performing stale models or high-performing enshittified models which know about more current information.
Indirect pollution via AI slop in the input and the same content manipulation mechanisms as SEO hacking is still a threat for open models.
A cautionary tale for many of these types of tech platforms, this one included.