The Downfall of DeviantArt
slate.com
slate.com
In no particular order, because I don't know which were profitable or which represented a larger portion of revenue:
- Subscriptions (users could pay for a few extra features and to disable ads on the site) - DeviantArt branded merch. - Prints and products with users' art printed on them - Sponsored Contests. These promoted movies or other media properties, or software of interest to artists. Often the prizes included Wacom tablets and Adobe Photoshop licenses.
During my time there a significant problem we were dealing with was due to deviantArt's stance on adult content. Anything was allowed as long as it wasn't outright pornography. In practice that meant that nudes were allowed but sexual acts were not. This had consequences for deviantArt's revenue. It meant that we could not run ads from the "reputable" ad networks and were forced to deal with seedier outfits that often (e.g. constantly) included malware in the display ads, exposing users to all sorts of nasty stuff. One of my projects was to detect and/or prevent the malware ads which proved challenging and at least given the amount of resources devoted to it, it was not very fruitful.
It really is sad for me to see what deviantArt has devolved into. Once the original founders sold out a few years ago I really didn't hold out much hope for the site's future.
I don't know the solution but isn't 4chan an example of what happens with no censorship? That would also kill an edgy contemporary art platform as the "4 teh lulz" crowd drowns out the "for the art" crowd.
As another semi-related example. Imgur used to be interesting "images" but now 25-75% are posts of text (like a screenshot of a tweet for example, or of a blog post, etc...). That might be good for Imgur's bottom line, I don't know, but it's not the site it started as.
To be the site it started as would require somehow disallowing images of text but magically allow meme images (some text). That might end up starting an arms race as users skew their messages onto billboards, tvs, signs, and other things to try to make them not look like just an image.
Like I said above, I have no idea if they want the site to be mostly images instead of mostly text. Only that it's another site that changed character over time and I'm sure didn't start as a site expecting mostly images of text.
In theory it's possible to have the same kind of site with the regular family friendly content and with strict (i.e. normal) levels of moderation. In practice we don't see this.
I think it's because the levels of active moderation by humans needs to be high. So 4chan can get away with less spent on moderation because their levels are lower.
That's how exactly how reddit mods operate, no?
Otherwise, reddit isn't a democracy, the mods can do what they want (within limits set by the admins), just like they can do what they want here.
Advertising: perhaps the best and most effective involuntary ally of the Puritan movement.
It is a shame that the Internet that we have created depends pretty much solely on that form of revenue.
The usual defense by advertisers is that their clients do not want their businesses to be associated with adult material, so they insist that advertisers do not place ads on adult-content-friendly websites to reduce the risk of that association happening due to ads adjacent to adult content.
I'm curious as to how valid that is. Are many/most businesses (or many/most businesses that are highly valuable to advertisers) materially concerned enough about association-with-adult-content-via-ad-placement that they would switch advertisers over it? Or are advertisers manufacturing family-friendly ad placement as a competitive selling point that their customers never really asked for?
I'm also curious about the next level down the stack, as it were: if a significant subset of businesses who hire advertisers are concerned about adult-content-association-via-nearby-placement, why is that? Is it merely disgust/discomfort on the part of business leadership and prevailing culture? Or does data indicate that there's a material risk to enough businesses' bottom lines that it's fiscally prudent to avoid that association?
(Reasoning from first principles without data, so probably wrong) I'm skeptical about the validity of both claims. Adult content is really popular across many demographics whose behavior is otherwise quite different. Given that broad base of popularity, is it really that risky for a business to have its brand appear in an ad next to pornography compared to, say, appearing on a politically extreme news site, or next to algorithmically-prioritized ragebait content on social media? Unlike adult content, I feel like the associations formed by seeing a company mentioned next to something anger inducing are more likely to be negative than the associations formed by seeing it next to adult content that the viewer presumably sought out. In both cases, the chance of behaviorally-significant associations being formed at all seems quite low, so I'd imagine that effects here (if there are any) would only be visible in the very very large.
Advertisers don’t have to believe the content is actually objectionable to fear reprisal. Outrage machine types on all sides will make an issue out of anything they can use to score points for themselves.
The marginal value of putting a Johnson & Johnson ad (say) in front of some furries is nothing compared to dread of the headline “J&J wants your children to be furries!”
A very real possibility when advertising on DeviantArt. But a no-adult-content policy would have (probably) satisfied adsense terms but would not have eliminated (most) of the furrie content on DeviantArt.
The advertisers care about getting bad press and having to do the rounds in the "news cycle".
By now everyone knows that advertisers don't choose which posts, tweets, images, etc. their ads are shown next to, but that still won't stop the CNNs of the world from writing a "<Brand> advertises next to Nazi content on <Platform>!" article.
Like, if a media outlet or politico wants to make hay about Walmart being advertised next to nazis, I’m sure that’s trivially easy to find today (and you only need to find one instance of that adjacency to make hay for a news cycle).
So why isn’t this already an issue in the status quo? Why aren’t brands being pressured into changing their ad placement habits right and left?
Sure, some businesses are being called out for things they explicitly endorse (e.g. Budweiser pride ads), but that’s not the same thing as running an ordinary ad in a questionable place.
Perhaps advertisers have written off adult content sites as places where there isn’t enough money to be made to risk it? If so, I’m curious what consumer behavioral data backs up that conclusion.
Didn’t this just happen to Xhitter? Their ad revenue has dropped like a brick.
That's the biggest joke in the world, the largest diffusers of explicit content are by far the advertisers which they use to sell anything, from cars to shampoo
The advertising industry is the very last industry which can claim a moral high ground on anything.
It's buck-passing all the way down.
Relevant Bill Hicks: https://www.youtube.com/watch?v=G4NyMJHWVHw
If the media stopped writing shrieking news articles about "X brand promoting Y evil", we could all move on collectively as a society.
I've worked in a couple of digital advertising roles (mea culpa) and brand safety is a very big deal. No-one wants their product associated with porn, extremism, or even politics. eXtwitter's revenue implosion due to advertisers pulling out (pun not entirely intended) is one recent example of the importance placed on brand safety.
In summary, they are a tiny % , but by no means insignificant in terms of impact.
Which is weird, given that "sex sells."
> If music be the food of love, play on; > Give me excess of it, that, surfeiting, > The appetite may sicken, and so die.
The point is actually to suggest sex. Keep you aroused, but not let you finish.
Imagine, if Budweiser ran a brothel and buying their beer actually got you laid. You'd get laid, forget about sex for a while, and the association between the beer and sex would not entice you to buy the product.
They're just leaving money on the table. Ads are sold at auction, so the auction price on NSFW pages would be lower, but because it would be lower then some advertisers would want the discount. The alternative is that the ad network bans them and gets nothing, the advertiser loses out on cheap impressions and the site goes bankrupt.
Sites like Deviant Art could have a policy to the effect of "NSFW content is allowed but you have to mark it" and then advertisers who don't want to be seen next to NSFW content, aren't. Hypothetically some user could post something NSFW without marking it, but the same is true on sites that ban NSFW content entirely.
Then the site itself gets the full payment on the majority of pages that are safe for work, still gets something from the ones that aren't, and doesn't have to deal with shady malware-laden ad networks.
Two main reasons:
- Websites cannot accurately classify content a lot of the time without human review. Ads are served dynamically, so they could appear on a NSFW post faster than a human can review it. It's less work to simply optimize for as little NSFW as possible.
- NSFW content consumers aren't valuable to advertisers or ad network operators because they don't sell products that could be placed as a 'contextual' ad. Ad platforms need insurance brokers, banks and CPG brands to make up the fat tail of revenue. That stuff can be placed next to posts about pretty much anything. NSFW content on the other hand doesn't work the same way, which is why ads on porn sites are for scam dating sites and unregulated erection pills.
This is doubly so considering the use of remarketing cookies. While you're watching something that's bound to fill you with regret moments after the fact, you won't want to see an ad showing the Amazon products you were browsing previously.
This is the same whether NSFW content is banned or not. "It's banned but somebody posted it anyway" has no apparent solutions different than the problem of "the user is required to mark it but they didn't."
As an obvious example, adult content is nominally banned on sites like TikTok, and yet there are thousands of TikTok accounts that serve as the funnel for OnlyFans, with the creators dancing on the line of the site's policy (and often crossing it), only to create a new account and carry on if they get banned.
> NSFW content consumers aren't valuable to advertisers or ad network operators because they don't sell products that could be placed as a 'contextual' ad.
Which is fine because the ad slots are sold at auction. The ad network's choices are to get something or get nothing.
edit: sibling comment has a better explanation which I think is probably 100% on the nose for this one.
It's a deliberate, active decision and its cost. That's policy. Prioritized over cost and profit.
In the US, there is often concern that the more valuable clients or end-users might otherwise boycot their product, or get them regulated even further. A concern that their are under pressure to do that, or else face even more costly consequences. (And they are. It's a very vocal part of society that's puritan.)
It's funny how significant percent of Facebook ads I see are an outright scam, phishing etc.
It's been like that for years.
It's App Store. Centralized censorship through App Store destroyed the Internet.
How does 4chan fund itself?
[0]https://www.polygon.com/23662953/good-smile-company-4chan-se...
Building DigitalOcean was fun, but building dA was 1000 times more fun, best times of my life for sure, very very grateful.
(I worked on gallery, help, irc, dAmn and a bunch of other stuff, 2001-2008 my emoji still exists :neom:)
Anyways, even though dA of the past is gone, it was definitely an important part to many people, so thanks again.
I wish the artists well in their AI copyright legal pursuits.
Whenever a platform is owned by shareholders who then need to extract rents from the ecosystem, this will happen. Whether it’s couchsurfing or twitter.
Expect it to happen to Reddit etc.
There is a direct line from the profit motive to platforms becoming enshittified, promoting the most outrageous content and making people emotional and angry.
The AI is just another level of appropriating human work. Whether it’s google’s disruption of publishers through AI-generated answers, or OpenAI training on artists’ work.
These platforms need money to survive. Automattic, the latest owner of Tumblr wrote a great post on all the things they've tried and how Tumblr is still losing $20MM a year IIRC.
Wordpress by that same Automattic doesn’t need money to survive, in the sense of money going to one large corporation. Anyone can self-host their own copy of wordpress, buy plugins etc.
If you want to know more about how to monetize digital content without a trusted central actor, we are working on a Web2 version of that ecosystem btw: https://qbix.com/ecosystem
Also, science and wikipedia and openstreetmap are examples of open gift economies.
Sorry the experimental stuff is not slick enough for you yet, we don’t have the resources of Facebook or even Automattic. We worked very hard for 12 years on the foundations at https://github.com/Qbix/Platform but I am sure you can find many faults there. (I’d like to hear about them btw.)
On the other hand, many other projects like the E programming language, Capn’proto, Linux etc. are also very complex and did not have fantastic and slick documentation, first adopters also had to read some words in order to get it.
This is an open source project. You are welcome to reduce the words and make a summary. Perhaps when we start marketing to a broad audience, we’ll reduce it to 10 word slides and sound bites, or jingles.
Until then you can try it yourself, the documentation is at https://community.qbix.com and technical documentation is at https://qbix.com/platform/guide
Wordpress' Tumblr problem is that costs exceed revenue, not distribution.
Note that "decentralized" doesn't reduce total costs - it just spreads them out. Server costs do not go down with the number of owners.
At a certain point, people seem to start looking at self destructive options to make that happen.
Social media sites that are still founder-owned or have strong individual leaders can continue fine (consider e.g. Dreamwidth). Though I guess whether you can sustain that past one person's lifetime is another question.
ActivityPub, used by Mastodon, Misskey, Lemmy, and Pixelfed among others is an example of such a protocol. BlueSky's ATProto is another, though it's in an earlier stage without mature third-party implementations and service providers. Email, too is decentralized, though it may serve as a cautionary tale; spam, attempts to block spam, and feature stagnation have all degraded the user experience considerably.
it's no wonder that the fediverse is most active on the fringes, especially outside of tech bubbles who use it because they like the idea
That's not to say a million fiefdoms is necessarily a bad approach. A small server where all the members know each other is much more likely to be run in a way that's satisfactory to all its members than a big one. Furthermore, users have the option to maintain multiple accounts.
And? I guess that somewhat hinders enshittification just by making it hard for the platform to ever evolve at all. But the cure is worse than the disease, you can't ever build something new that way nor can you really improve something that has any level of traction. Look at how IRC users revolt when you try to fix even the most glaring problems.
ActivityPub dictates how to meditate relationships between activities on collections of objects, and a few default objects types. It's specifically designed to let you "subclass" (not really inheritance, and more composition as you give a list of types with no enforced hierarchy) objects so you can create new object types nobody else understands but still give them another type that allows them to carry out basic operations on it.
It's not perfect but it's far from as dire as you make out.
Enshittification is hindered not because nobody could create a Mastodon fork (or green-field project speaking the same protocol) that's riddled with ads, but because people can select a different service provider and still access the same network.
In theory, certainly.
In practice, one instance dominates and all the other instances have to censor accordingly or die.
Why not start a centralized social-network service as a non-profit / benefit corporation, paid for by donation?
A social network has to make money somehow because it has bills to pay. Hosts aren't free. Servers aren't free. The cloud isn't free. Staff isn't free. Moderators are usually free but shouldn't be.
Sure, but probably a FOSS social network would need far fewer of these than a paid one, because 99% of the server costs of something like Facebook or Twitter, go toward the backends, analytical DBs, and graphical / ML models used to power "features" that no user wants, but which make Facebook themselves money.
And a FOSS social network would just... not build those kinds of features.
Instead of, say, a feed constantly rebuilt to drive engagement and rage-bait, you'd just get a simple chronological feed of what everyone you're following is posting; with maybe the ability to dial down the number of posts you see from any given person you're following ("show me the only the most-liked Nth percentile of posts from this user") without unfollowing. But even that kind of filtering — in fact, even the merging of followed users' feeds! — could all be done client-side. The whole "feeds" part could be as simple as a post-triggered static-site-generator pushing JSON into an object-storage bucket hiding behind an edge cache.
And I may be wrong but I don't think the recommendation algorithms and other such features take up as much of the cost as you're claiming. I think a lot of the cost of something like Facebook is probably taken up by infrastructure and storage. Recommendation algorithms probably aren't that expensive.
Having extra entire complete copies of the relationship-plus-posts graph, denormalized in various ways (incl. in ways that inherently prevent use of easily map-reducible algorithms, and so require heavy vertical scaling) such that you can run the algorithms, is what’s expensive.
And constantly feeding the data into those denormalized models, using specially-tuned realtime ETL technologies that themselves do distributed scaling to ensure no infinite queue backlogging from activity bursts, is also expensive.
I don’t think it’s true that it’s intrinsically impossible for a public service to be self funding, and I think that not everything has to grow / change forever to remain relevant.
We need to figure this kind of stuff out, I mean Wikipedia is nice and all but it’s really bad that humanity in general has to rely on megacorp for things as basic as maps while we say we’re living in the Information Age
There's other AP implementations that aren't a constant server hog like Mastodon is and can run on much weaker hardware (some of it can run on a raspberry pi). You don't need a full rails stack if your user count never exceeds 100 (which y'know, is the ideal state of AP - small communities who can remotely interact with each other).
Modern Usenet :-)
This is incorrect. From direct personal experience, I strongly believe the backend infrastructure and staff required just to operate the core product functionality of a successful large-scale social network massively exceeds what could be provided by donations.
Just in terms of core product OLTP data, Tumblr hit 100 billion distinct rows on MySQL masters back in Oct 2012. At the time, after accounting for HA/replication, that required over 200 expensive beefy database servers. This db server number grew by ~10 servers per month, because Tumblr was getting 60-75 million posts/day at this time.
Then add in a couple hundred cache and async queue servers, and over a thousand web servers. And employees to operate all this, although we kept it quite lean compared to other social networks.
Again, this was all just the core product, not analytics or ML or anything like that. These numbers also don't include image/media storage or serving, which was a substantial additional cost.
Although Tumblr had some mainstream success at that time, it was still more niche than some of the larger social networks. At that time, Facebook was more than 2 orders of magnitude larger than Tumblr.
Because these social networks are designed for analytics. It's in their blood. It permeates everything they do, and causes immense overhead.
Check out mailing lists or usenet.
Social networks store a lot of OLTP data just to function. Every user, post, comment, follow/friendship relation, like/favorite/interaction, media metadata -- that all gets stored in sharded relational databases and retrieved in order for the product to operate at all. For successful social networks, it adds up to trillions of rows of data (on the smaller end, for something like Tumblr) and that requires a lot of expensive infrastructure to operate. Again, none of this has any relation to analytics.
As for usenet, what? It's basically dead, after becoming an unmanageable cesspool of spam (or worse) more than two decades ago. It was great in the 90s, but the internet population was substantially smaller then.
Yeah, because no one is interested in promoting it because it doesn't have analytics baked in so you can't make money from doing so. Of course it deteriorated over the years. It's also cheap to run and can handle a massive amount of users.
> store a lot of OLTP data just to function
Right, so they can run analytics. You could reduce your tracking data to aggregates, but then you can't go back and run analytics on your users. You don't need to keep that data forever.
Especially with modern social media where content older than a day is effectively dead and ignored.
> it adds up to trillions of rows of data
This was a lot of data a decade ago. Nowadays a single postgres instance will handle billions of rows without breaking a sweat, and social media content is exceptionally shardable.
Stop gaslighting me, it's not OK! I'm describing first-hand experience of things that were not related to analytics IN ANY WAY, SHAPE, OR FORM.
Try running OLAP queries on a massively sharded MySQL 5.1 deployment, or any aggregation at all on a Memcached cluster. These technologies were designed for OLTP data, and were woefully incapable of useful analytics over massive data sets.
I was Tumblr's fourth full-time software engineering hire. When I joined (nearly 4 years after the company was founded) the only thing remotely related to analytics was a tiny Hadoop cluster, where logs were dumped and largely ignored. Nothing about analytics is "in their blood". All you needed to sign up for Tumblr was an email address. WTF do you even think they are "analyzing"? Your comments are completely fabricated BS.
> You could reduce your tracking data to aggregates
Once again, I'm not describing "tracking data"! I'm talking about things like content that users have posted, comments they have written, content they have favorited, users they are following. These are core data models of a social network. It has nothing to do with tracking or analytics.
> You don't need to keep that data forever.
The OLTP product data I'm describing does need to be kept forever. Users don't like it when content they have written on their blog suddenly disappears.
> Nowadays a single postgres instance will handle billions of rows without breaking a sweat, and social media content is exceptionally shardable.
Yes, but running a massive cluster of hundreds or thousands of sharded database servers is still very expensive.
There is no central concept of "content that have favourited" or "users they are following", that's all handled locally in that model.
Not by volume. Posts and comments make up the vast majority of the storage requirements, and none of that can be purely client-side.
> There is no central concept of "content that have favourited" or "users they are following"
I'm aware, I used Usenet quite a bit in the 90s, as well as dial-up BBSs.
Usenet is a distributed forum / discussion board, which is related but not equivalent to the core functionality of social media applications being discussed here.
With Usenet's model, there's no concept of a profile aggregating content from a single user. This means you simply cannot replicate the primary experience of Facebook, Twitter, Instagram, Tumblr, Pinterest, DeviantArt, MySpace, Friendster, or any other social media site/app with Usenet's approach. Nor can it reproduce the experience of even modern forums like HN or Reddit.
Usenet also didn't actually scale massively. Every estimate I've seen of the peak Usenet userbase puts it at a tiny fraction of modern social media.
In any case, Usenet essentially failed. We already have empirical evidence about how these ideas play out! Why are we even seriously discussing this?
These aren't really good "social media" examples. Both mailing lists and Usenet have limited retention, with mailing lists there may be almost no retention beyond the amount required to deliver a message.
While low retention might be a desirable feature and something you might actually want in a FOSS social network, it means old content will disappear from the central server. If it's not archived by clients it can easily disappear or end up locked away only in private backups. Google's buyout of Deja News should be a cautionary tale of retention and the locking up of public data behind a private gate.
Usenet history today is largely only available because someone at Google hasn't noticed Google Groups still exists and terminated it yet. If that happens tomorrow there's not any good complete archive of historical Usenet content. There's no guarantee Google won't kill those Usenet archives in the next year let alone the next five years.
We have 1.5 trillion distinct rows living on a PG cluster right now. Growing at a higher rate. And that requires... four DB servers (four shards). And no replicas — all our prod query traffic hits these servers directly, as a mixed workload with the writes from our queue-consumer ETL agents.
(The servers are "beefy", sure — but they're not that beefy. They've got 128 cores and 1TB of RAM each. That's not even top-of-the-line these days.)
Despite serving a whole ton of QPS, these DB servers are not CPU-bound or memory-bound or even IO-bound. Our primary horizontal-sharding factor, in fact, is PCIe lanes — we're constrained mainly by the number of NVMe drives we can keep connected performantly per CPU [1] — and thereby, generationally, by "fast storage capacity" per server platform.
One thing that's perhaps "special" about our use-case, though, is that our data is inherently append-only. Which is pretty great for OLTP performance: there's very little write contention, as we just partition the data by insert time (which is also a natural partitioning key[2] for most queries our business layer does) and then only write to the newest partition, with all previous partitions becoming effectively immutable.
Most workloads aren't like this... but they could be, if you model your data carefully, with temporal tables holding versioned records and so forth. You trade off more storage growth, for far less worry about write contention. ("Event streaming" is not the silver-bullet solution it would seem here — as you would still need to create an aggregate to query from; and writing into that aggregate would still be an OLTP write bottleneck. No, the true solution here is data-driven [schema] design — i.e. forcing your API engineers to invite the DBA to feature design meetings, and ensuring the DBA has the perogative and incentives to say "no" to a bad design. Or you could just hire database systems engineers to code your business layer, like we did.)
My point here, is that if you're very careful in designing for not just "scalability" but economies of scale — and if your org is engineering-driven rather than product-driven (as a non-profit social network would likely be!), and so has "mechanical sympathy" — then you can achieve things with just the budget of an average B2B SaaS, that would rival a (smaller) social network.
But with the budget of a benefit corp that nevertheless charges even $1 for each install of its social-network mobile app? Who knows what you could achieve!
(See also: WhatsApp, pre-acquisition. They were just charging $1 for each install. With just 50 employees, that got them a lot of operating budget.)
---
[1] And actually, that's part of why the DBs are not IO-bound. Put enough NVMe disks in RAID0 together, and you get something nearly functionally indistinguishable from memory. (Yes, RAID0 — because our DBs aren't the store-of-record; our data lake is.)
And even where NVMe reads are not indistinguishable from memory reads, they're still complementary. Postgres, in relying on mmap(2)ing heap files and thereby on the the OS page-cache as a buffer pool, makes per-query serial page faults expensive, and causes multiple threads wanting the same cold data to bottleneck by repeatedly becoming coalesced awaiters on the same sequence of pages faulting in. But when you've got highly-concurrent queries [= many different OS threads] that are all page-faulting for different cold data, then you get deep IO queuing, and everything works out optimally. So the expensive cases don't really come up — the page-cache serves the head of the request distribution, while the deep IO queues of the RAID0-NVMe-pool are a perfect match for the long tail.
This makes it almost irrelevant, at runtime, whether data is hot or cold. As long as you have sufficient memory to host the very hottest data (e.g. relatively-low-cardinality fact tables like users/blogs that get joined to everything, and esp. their indices), the rest can be read straight from disk with nearly no penalty.
---
[2] Funny enough, given that we're doing heavy joins for some queries, the one thing we do sometimes run out of at runtime, is timeslices of the Postgres postmaster to coordinate lightweight locking of the global locks table.
We not only have time-partitioned tables, but also something akin to tenant-partitioned schemas. This results in a lot of relations. (It takes hours for us to run vacuumdb --analyze-only.)
And some OLAP-ier queries (that we nevertheless have to run, synchronously, in response to client requests!) need to touch many of the partitions and many of the schemas. That's O(MN) locks they need to take — which take time to write into the global locks table, even though those are only memory writes.
Every once in a while, we have to "consolidate" our table partitions — not by rolling them up with aggregations to reduce row-count, but rather by just copying all the data into fewer, larger partitions — so that we can take fewer access-shared locks per query, so that transactions can spend less time waiting on a handle into the global locks table at startup just to write these O(MN) read locks into it. (In other words, we consolidate data to reduce O(MN) read locks per transaction, to O(N log M) read locks per transaction.)
But that's a PG problem, not a resource problem per se. We've been considering writing a patch for PG to allow relations to be marked as "immutable", where an "immutable" relation doesn't have locks tracked for it (or its indices) at all — but also can't be written to, re-indexed, or even dropped. We'd then just apply this setting to all of our historical partitions. (If you're curious: given its semantics, this property could be enabled for a relation at runtime; but would need an instance restart for disabling it to take effect, as the instance would have no idea what read-locks "would be being held if the table would have been tracking locks", and so couldn't safely do any writes to the table without a barrier that drops all existing MVCC read-states — i.e. an instance restart.) As a bonus, such relations could also have "perfect statistics" calculated for them by ANALYZE; and those stats could then be marked as never needing to be re-calculated — allowing both VACUUM and ANALYZE to be skipped for that relation forevermore.
"Four shards" and "no replicas" is a complete non-starter for a popular social network. That low shard count means a hardware failure on a shard master results in downtime for 25% of your users. Putting ~400 billion rows on a single shard also means backups, restores, schema changes, db cloning, etc all can take a tremendous amount of time. No read replicas means you can't geographically distribute your reads to be closer to users.
I'm not sure why you're assuming that social networks aren't 'very careful in designing for not just "scalability" but economies of scale'. I can assure you, from first-hand experience as both a former member of Facebook's database team and the founding member of Tumblr's database team, that you have a lot of incorrect assumptions here about how social networks design and plan their databases.
I can't comment on any of the PG-specific stuff as I don't work with PG.
Append-only design is not usable today for social networks / user-generated content. That approach is literally illegal in the EU. In years past, afaik Twitter operated this way but I was always skeptical of the cost/benefit math for their approach.
The "200 expensive beefy database servers" I was describing were relative to 2012. By modern standards, they were each a tiny fraction of the size of the hardware you just described. That server count also included replicas, both read-replicas and HA/standby ones (the latter because we did not use online cloning at the time).
I haven't said anything about event sourcing as it generally is not used by social networks, except maybe LinkedIn.
Didn't we have some more forms of government available for discussion?
E.g., after the xz backdoor, I read a call of OSS maintainers that critical OSS projects should be state-funded as they literally comprise critical infrastructure. Why couldn't we do the same with social networks?
If the goal is to avoid "enshittification" I don't think the solution is to have governments control social media and then have those platforms be subject to bureaucracy, decency and profanity laws, surveillance and propaganda (more so than currently) and the fickle wrath of the taxpaying voter. Realize how little critical infrastructure actually gets funded, and then add the psychotic dog-water paranoia of half the US thinking that infrastructure is a psyop by communists and "groomers" to turn their kids gay, and making that an issue in the polls.
PBS is probably the closest analogue to government run social media I can think of. Half the government and their constituents consider it "liberal propaganda" and want to defund it entirely, and it has constantly has to go begging hat in hand just to stay afloat, and then pursue commercialism to make up for the deficit.
You think social media is bad now? Imagine if you're required by law to sign up with proof of citizenship and your SSN is your password. And every platform is constantly putting Wikipedia style donation popups. And it's a misdemeanor to post swear words or material deemed "inappropriate."
Functionally, the “public good” part of social media networks is almost certainly better served by a single organization.
However, the “freedom of speech and ideas” part, runs in horror at this idea. (rightfully so).
The best middle ground concept I’ve heard is to contrast the current state of the web with libraries.
NB: Enshittification is going to become a term like “Fake news”, completely divorced from its original roots.
I would expect government run social media to be especially enshittified, honestly
Just for different reasons
There are countries which CAN make it work, but man, a central nervous system co-opted by oligarchs, tyrants or other worst case scenarios, would be the outcome for most of the world.
In contrast, DeviantArt saw dollar signs from AI art and rugpulled themselves. Their business model relies on art remaining scarce enough to not exhaust the demand for art. A machine that lets you create unlimited art for the cost of some GPU time completely destroys the economic underpinnings of most artistic endeavor. While not all artists are solely economically motivated, the ones that are economically successful are the ones paying for dA subscriptions - the things that keep the site alive.
Before, under Yahoo, they'd just put some new hoop in the app to prevent users from accessing adult content, but by the time this particular Apple review rolled around, the sale of Tumblr to Verizon had already been finalized. Which created a situation where an outgoing management pretty much ordered to not bother fixing it and just banning it all, hoping that Verizon's "family friendly" policies meant that it wouldn't jeopardize the sale (which it turns out, the sale wasn't jeopardized).
You can still kinda see it in how broken the actual removal was; they just excised the NSFW from the frontend by marking the posts as sensitive on the API and then preventing the frontend from viewing anything sensitive. For years (and maybe even today) you could just scrape the API to find NSFW posts, although that's on the decline since most NSFW Tumblr accounts have been deleted entirely by the actual people behind them.
That's not true - Tumblr already had their "no adult content" plans in place well before the CSAM problem caused Apple to temporarily suspend them from the app store. That suspension just brought the plans forward by 6 months in a panic rush.
And it's not just Apple, payment processors also have strong opinions.
In general, sexual/erotic stuff has become a really hard thing to keep allowing in mainstream platforms.
Using that situation as a reference point, I'm not totally convinced that AI art will supplant digital art/drawing. AI art can be used to produce some types of drawings, sure, but not all types of drawings can be convincingly produced using AI. AI's stylistic and compositional abilities are constrained to its training set, for example.
Another thought: when the DSLR became affordable in the early 2000s, it wiped out an entire segment of low-level professional photographers. Suddenly, any schmuck with a DSLR could take "good enough" photos with ease, and without the friction of film. I'd expect AI art to have a similar effect in the realm of digital illustrators. But, professional photographers still exist, and so will professional illustrators.
Even categories that are supposed to be for specific mediums where AI shouldn't be applicable are full of it regardless - just now I scanned the Photography section and almost immediately spotted a conspicuously three-fingered woman. Posts made using AI are supposed to be tagged as such so users can opt-out of seeing them, but that "photo" isn't tagged, nor is anything else on the uploaders profile despite all of it being blatant AI.
You could almost turn it into a game - pick a random category and see how far you have to scroll before you see anything at all that doesn't scream "babbies first copy-pasted MidJourney prompt".
While this is the official line I'm fairly certain the real reason for stuff like this is to prevent AI models consuming their own output.
Now, with Wix in charge, and a handful of the roachier staff left (I'll name names, Realitysquared) the site has negative bupkes chance of content moderation/curation worth a wet damn.
I actually started writing a long screed about this on my own blog last year with the AI debacle but shelved it (now I feel compelled to return to it)
DA had plenty of small problems and a few big ones, exacerbated by incredibly green leadership at the top. It was pretty much run off the whims of the CEO, who even when he had good ideas, had no execution focus, and often outright ignored capable feedback from his lieutenants.
Things like having an app, an API, a GTM strategy that was remotely baked for feature releases was simply non-existent at BEST, and done in the most haphazard more typically.
DA eventually shredded its own community; exhausting most of its most dedicated members, and ignored offers for acquisition that made much more fiscal and practical sense then the pittance it eventually went for to that trashfire place called Wix, who in all likelihood will eventually sell off bits for scrap.
DA was a true gem for a decade, and then fell in spite of that due to categorical poor management and vision (or the execution thereof).
So yeah, I have no disagreements with you there, but imagine what they could have done with a real product roadmap that didn't change on a whim?
>How is this democracy?
If you are in a democratic country, you can vote for representatives, or run for office yourself. Has nothing to do with you being able to host something on someone else's computer.
>How is this an advanced society?
Because you can upload functionally limitless digital content to someone else's computer and someone else halfway around the world can almost instantly view it on a handheld device. Obviously restricted to the rules of the entity paying to host said content.
This retort always seems kind of trite and disingenuous.
First of all, no, they are not free to do these things unless they are billionaires with capital to spend. "Just start your own business" is a silly dismissal because the vast majority of people cannot start such a business. Second, it doesn't even address the original complaint. "I think Company X should do ABC." Suggesting they start a different company to do ABC doesn't really address what OP calls for: Company X changing their ways.
There has never been a cheaper and quicker way to share one’s art in the history of humanity.
It's just that the network effect of megacorp social media means you won't actually get many random arbitrary visitors and you won't have a way to monetize easily.
So if you just care about art and not about monetizing you're art there's no problem. Should be good. If your goal is only money then, yes, you have to stick to the marketplaces.
I remember being at the gym in England in the middle of the day with my American wife and the TVs were showing a BBC show about body positivity which basically involved full frontal nudity of all the participants. She was very surprised. To me, it was just Tuesday.
Another time in Switzerland I was watching an evening game show and the host inexplicably undressed. Here's a pic seconds before the real action: https://imgur.com/a/n3iLhps
Europe isn't for beginners lol
The reason you can't post an anatomy tutorial is unrelated though. That's because YouTube is cheap. They don't want to spend the effort to try to differentiate an art tutorial from more problematic content.
Professional digital artists post at Artstation now from what I can tell.
You sound like someone who missed the word deviant in that website's name.
Care to share a link to the art you make?
Really curious about what kind of sophistication you're producing that would be tarnished by being seen next to furry scribbles.
Great. Give yourself a pat on the back for that one.
Now that you've had your smug moment, note that the parent comment wasn't concerned with the quality of content, but the category.
"Furry smut isn't art", brought to you by "Rap isn't music" people.
There's still people on there selling good reference material but now they're buried between a few dozen variants of 6000+ WIZARD CHARACTERS IN 4K!! that someone churned out with Stable Diffusion in an afternoon.
This used to be an entertaining experience to see all the kinds of beautiful art people used to create... But as you said it is getting filled by AI crap. I noticed the same thing with Pixiv, even if I opt-out from AI art, new results from several popular tags are filled with badly tagged AI images.
Now the experience feels tiresome, in some categories I waste too much time filtering between real and AI art. I started to search things produced before 2022-2023 to avoid AI.
This "art" should be uploaded to its own subdomain and never get mixed up with real art.
It's a spam problem, only worse, because they're actively paying the spammers.
Even Spotify has this problem. All too often I'm getting recommended crappy remixes "slowed and reverbed" or "sped up". Just recently I got some guy's crappy techno with the artist field spammed with completely unrelated bands I follow. Of course, when I tried to report this, Spotify only cared if the guy was selling bootleg merchandise.
The whole thing made me click, "hide artist" and "hide song" for the very first time.
DeviantArt built a system that pretends bots are not a problem on the Internet...and gave them a profit motive.
https://www.webdesignmuseum.org/web-design-history/deviantar...
And 2012, which I consider peak DeviantArt:
oh well, it's not 2012 anymore, let the past die.
You needed a GitHub account and a KeyBase account. So people created as many accounts as their bot networks were capable of, and tried to get the crypto.
Thankfully KeyBase changed the requirements to include "account must be X weeks old".
Edited to add: I'm not sure if there's a way to prevent bots these days. Feels to me that we're lucky (more?) economic systems haven't been bled dry by bot networks.
I miss the promise of KeyBase. It felt like a real digital identity, but for whatever reasons it wasn't good enough to succeed.
I am sure that LLMs and bots will be able to fool many people on HN and run “rings” around dang’s ring detection software, in about 5 years. It’s a gameable metric, after all.
They were already able to do it on 4chan in 2020 with just GPT3! And the most impactful thing is users started accusing each other of being bots! It literally enshittified the whole forum overnight:
https://finance.yahoo.com/news/breaches-every-principle-huma...
And manually confirm with the companies at random.
They’ll just check in and then run bots in their account. Line a chess bot for example
The reason, imo, was the acquisition by Zoom and apparent total abandonment of the project.
That was the stumble that gave room for other platforms to grab pieces of its then-current and future user-base. Anyone can tell you that a very large portion of the money changing hands online for art (adult and not) is actually changing paws, so dA missed out of having a slice of that, whether through advertising or facilitating transactions. Worse, its reputation was tarnished among adjacent subcultures.
There have also been regular ToS panics every 2 or 3 years, where someone's (mis)interpretation of the licensing rights dA claimed for being able to modify and distribute artwork (i.e., make thumbnails and send images in daily update emails) caused users to swear off the site for fear of having their work "stolen". Add to that, quite a backlash against the recent site redesign (and the ones before it).
That is all to say, this really has been a long time coming. My account is nearing the 2-decade mark, but I haven't logged on more than a couple dozen times in the last half of that. There's just almost nothing there you can't find more easily or comfortably elsewhere.
- hub tries to monetize its audience, rightly so for sustainability of the platform. Server costs, management, moderation etc etc.
- it takes just a few years for hub to become a corporation full of clueless corpos who have no idea about the initial culture and core audience.
- hub is bleached beyond recognition because corpos are scared of anything that is slightly controversial - including the original culture hub was about.
- "That's not family friendly! We also need those ESG and DEI labels to attract investors! The advertisers won't approve their brand associated with this!”
The current term for this is "enshitification" right?
- hub dies, culture might disperse, corpos get their golden parachute and latches into the next big project.
It seems to be a very common story. Reddit, DA, you can think go a lot of examples. It WILL happen with big Mastodon and BlueSky instances.
If you think about it you gotta give credit for 4chan keeping a big chunk of its soul, rather you like the site contents and its users or not.
1) ESG/DEI stuff wasn't even a whisper on the wind when this stuff went down. It was still very much DADT, Jack Thompson lawsuits, elected-Bush-twice America. In other words, the complaints were mostly coming from the center/right (and Joe Leiberman, if you still considered him a liberal).
2) 4chan was in the "advantageous" position of never being able to attract major advertisers in the first place. And while a good bit of the old culture is still extant, the post-Trayvon-murder/GG/Trump-meme-magic era did a number on its userbase's ability to focus on the lulz instead of descending into conservative (if not just straight-up Nazi) rhetoric (and not for laughs).
But, yeah, the rest of this tracks. It's basically inevitable unless the site admins get to a point where they're happy with the userbase size/culture/whatever and decide there's no need for any more changes. Examples: Craigslist, SA (to a degree), FA (despite controversies and the recent UI change, which users can mercifully opt out of). In fact, I would say that unnecessary or large-scale UI changes are good heuristic for determining when things are about to go downhill.
In the same way that Playboy can get away with posting sneaky things and Meta can only shadowban them and doesn't dare terminate them: https://www.instagram.com/p/C3-nnn9RM82/ (NSFW)
Not saying it doesn’t suck but ALL websites have an AI spam problem. Twitter, Pixiv, ArtStation etc are no exception.
Since the aptly named "Eclipse" redesign it became terrible to use, so I stopped.
The lean into AI will just let it rot for a few more years than otherwise.
Its a shame to see what its become, the ascension of AI slop means a site like it will never be possible again unless there is some incredible filtering capabilities available.
In DeviantArt's glory days, 50% of the content was derivative fan art. Machines are pretty damn good at making things that already exist.
That's not a direct contributor to the demise of an image sharing site, no matter how much DeviantArt dresses itself up as a Web 2.0 era hub. "It's like a virtual silk road specifically for artists all over the world!", wonder how long that can stand in the age of monolithic social media platforms. Sites which are more than capable of serving what DeviantArt specializes in.
Sent over what I wanted, some initial payment... and then around 1 week later, they showed some hideous, horrible, amateurish logos that had nothing in common with those that were promoted on their page. Of course my money was taken, some even told me I can't do anything with them. And that was true.
Well, it took them a few decades to implode...
A number of years ago, someone posted an outstanding wireframe that I wanted to license. I was willing to pay quite well. I had just started a company, and the wireframe would have been quite useful in branding. I probably would have contracted the artist to do the rendering, as well.
I found it pretty much impossible to initiate contact with the artist. I think I got through, but I have no idea if I was successful, as they never got back to me.
It was mostly social media, really.
If I'm a painter and I look at a picture of a drawing I'm going to catch some of the style, some of the color palette, some of everything. My next painting will inevitably have some influences from that drawing .. and the other thousands I've seen. Why is it because it's now done in binary is it any different?
The second issue is that AI does not learn like humans do. Two different people looking at the same image will not be inspired the same way, because humans are inspired subjectively. There is a layer between what you see and what you learn, because everyone is wired differently, there is creativity involved in this process. AIs are not like this, they are not creative and there is no subjectiveness, it is built on plagiarism.
So yes, I agree on not treating binary differently. When a human learns from other artists and applies it without any creative twist, we call it plagiarism. When an AI does it, we give it a pat on the back and say it's just learning from what it sees. Why should AI be treated any different?
When I was a student, I got my eye candy from the school library, which provided context and history to the art. On these site, most images are presented without much context: what was this artist influenced by? How did they develop? Etc.
Only when the crowd effect of "AI=bad" came along, the pitchforks came out. As someone to whom AI art has brought incredible joy, it's very disheartening to see artists and the public straight up refuse to understand both the technology and the artistic potential -- the human side of AI art.
By using a template? Because based on my experience with inDesign and Affinity Publisher, it's still required to have knowledge about design and typesetting. They reduce the costs to get started and work in the domain, but the knowledge requirement was still there. Same with digital drawing and photo retouching. You're no longer gate-kept by the material costs. But AI is the equivalent of pressing X in a fight game and then saying you can do MMA and ready to go against UFC champions.
However, getting the picture you want, consistently, is a little bit more work. You might need to involve many more tools, including some old ones (blender for setting up, photoshop or gimp for post-processing) , and some weird new ones (like what the heck is a LoRA?)
It's like the one time I wrote an essay using LaTeX. Even when I was half-way done. It looked really well typeset and professional from the get-go, but of course half of the text was missing and still needed to be added.
Well, "most people" might be inaccurate. By volume, the people causing AI-generated content to come into existence are overwhelmingly just cranking the handle to churn out content, and that volume overwhelms everything else to the extent that it appears to be "most people", but might only be a few hundred. In the time it takes you to produce one artwork, they've got ten thousand 4096×4096 squares.
It only came along after a crowd of "AI artists" who are just spammers came to the platform. I mean every platform.
Really?
> Only when the crowd effect of "AI=bad" came along, the pitchforks came out.
You don't see how you spamming "a lot of" work onto a site reduces the value of that site? What you call the "AI=bad crowd", others were calling the "anti-spam" crowd.
I'm not sure how you would characterize this as spam.