AFAIK the original reason to close old threads was comment spam but that seems like it's a somewhat solved problem or solved enough it's no longer a valid reason to close threads.
AFAIK the original reason to close old threads was comment spam but that seems like it's a somewhat solved problem or solved enough it's no longer a valid reason to close threads.
Question and Answer is top Google result due to well formulated Q, keywords, Google algo, whatever. OK, start to see interesting responses. Closed as Off Topic or Duplicate by randos.
I argued, how can it possibly be a duplicate or off topic if it's a top search result. You're rewarding overzealous types. I suggested drastically increasing the points required to mark posts like that. I doubt I influenced anything, but it was worth a shot.
I also think a technical reason due to the scale of reddit, etc. I see no reason to do it other than what I mentioned on discourse, phpbb, vbulletin, etc.
I don't see the problem with "omg me too!" on an old thread as long as "omg me too!" is acceptable on a new one.
Just last week someone posted a question on a thread from over a decade ago asking what the end result of building something was.
I and several other people party to the original discussion explained the result, the performance of the system over the past decade and that options were different now
Let's see Reddit do that.
Bumping threads from a long time ago is a feature. Not a bug.
Comments could be marked by mods (or voted on) to not bump a thread, making it a minor occasional annoyance rather than a real problem.
> I also think a technical reason due to the scale of reddit
Reddit also just stopped working on UI (other than to make it worse) a long time ago. Why not have an additional view to "new", "top" etc. that sorts threads as if new comments bumped them?
That one should only talk about what a lot of other people currently want to talk about, or only for a certain amount of time, is really a dead end to me.
The tiny tin-pot hobby forum can keep every single post in memory and it's not really a problem. They can also do a full database scan pretty quickly.
But reddit can't do that. If it's not in the cache, it takes a lot of work to pull the data from the database. And there is no way to cache the entire dataset. That's why threads get locked at 6 months. So they can be statically archived for quick access.
Economies of scale isn't really part of it, it's more about moore's law.
When Google+ was shutting down, I did estimating of the total amount of public content within the Communities feature. Median size of a post was pretty close to Tweet-sized --- 120 characters or so. (G+ could ingest very large posts --- I never hit a limit though I wouldn't be surprised if it was book sized.) The highly active user population was maybe 12 million (another 300m or so posted at least once). Volume seems to have been about a million posts per week, for six years.
Which works out to less than 100GB of actual user-contributed text.
https://old.reddit.com/r/plexodus/comments/afbnvd/so_how_big...
Adding in the rest of G+ multiplied that out a few times, but likely still well under 1TB data. Images were asssociated with about 1/3 of posts and weighed in at a few MB each, call it 3 MB --- so a picture is worth 24,000 posts.
Rendered and delivered, page weight was just under 1 MB (excluding graphics), for a payload-to-page ratio of 0.015%. (Ironically, about the same as the ratio of monthly posting users to all accounts.)
But if you wanted to extract and store just user content and metadata, excluding video, storage requirements are surprisingly modest.
Facebook data aren't clear, but:
> There are 2.375bn billion monthly active users (as of Q3 2018).
> In a month, the average user likes 10 posts, makes 4 comments, and clicks on 8 ads.
> Hive is Facebook’s data warehouse, with 300 petabytes of data.
> Facebook generates 4 new petabytes of data per day.
https://www.brandwatch.com/blog/facebook-statistics/
I'm going to assume much of the 4 PB is system data, not user-generated.
2.4 billion users posting 4 x 120 byte comments/mo works out to ~15 TB of textual data per year.
Values are extrapolated, unconfirmed. Corrections or suggestions welcomed.
New comments don't need to live in the same read-only storage as the old ones. They just need to be found when the old ones are rendered, and that's easy - the new comments are in hot storage after all.
It's really sad that some great threads aren't allowed to live.