Granted, I'm mostly working in small-to-medium codebases, 20k-30k LOC incl test suite. I wonder if that's a factor in my positive experience. Curious to hear your thoughts.
31 karma · joined May 6, 2019
Granted, I'm mostly working in small-to-medium codebases, 20k-30k LOC incl test suite. I wonder if that's a factor in my positive experience. Curious to hear your thoughts.
The history of computing is full of predictions that consumer hardware would catch up to server-class capability in X years, and the answer has consistently been, consumer hardware catches up to _yesterday's_ server capability while server capability has moved on to new more mind-blowing paradigms which would not be possible on consumer hardware for another half-decade or more.
I'm sure that specific scaling trajectories will hit specific ceilings, such that in specific ways, one can make the argument that (for example) today's iphone performs at parity with today's servers. In 5 minutes I can spin up the same Postgres or Mongo DB that the largest companies on earth use server-side, though I can't support anywhere near the same data & traffic volume. But parity along specific technical aspects is a very different matter from the broad prediction of "you won't need a server for SOTA".
To step back to the bigger context -- your original point seems more along the lines of "we're obviously in an unsustainable bubble, and the rapid progress in on-device AI will further exacerbate the embarrassing collapse of all these overhyped AI companies". I strongly agree with you. But I think that's likely _and also_ firmly predict that the technical SOTA of 2031 (and 2041, if we make it there), in nearly every imaginable aspect including language-capable AI, will be vastly more capable than what you can run in your pocket.
Naw man, you crazy. If you tell me that in 5 years, consumer chips will be so good that I can run GPT-5.4-level AI on my phone, I'd find that plausible (I buy cheap phones). If you're telling me that in 5 years we won't need _servers_ because our _phones and/or desktops_ will be powerful enough to run the biggest newest LLMs in existence, I question your judgment, I think that prediction shows a deep uncreativity about how massively compute-hungry SOTA models will get.
The valuable things to do with inference will keep being a server niche because they'll keep being 1-2 OOM more compute-hungry than whatever consumer hardware can handle. Like gaming: my laptop can run games from 2015 at max settings no problem but the games actually worth getting excited about in 2026 still melt a $2k GPU, because whatever headroom the hardware gains, developers immediately spend on ray tracing and Nanite and modelling individual skin cells or whatever. I don't see any plausible reason to expect that the ceiling on "valuable server-side compute" or "inference capacity" will rise any more slowly than the on-device capability is rising.
My assumption is that in 2031, SOTA top-intelligence AI will be hosted on cloud servers like it is today, offering dirt-cheap access to capabilities we can't even dream of today, while your Android will be running some open-source GPT-5+ equivalent.
If you're listening to Spotify autoplays and a shitty song comes up, skip it. If AI slop is flooding Spotify with shitty songs, they'll naturally fail algorithmically (assuming we trust Spotify to actually be honest about its algos, which I'll admit we shouldn't https://substack.com/@tedgioia/note/c-236242253)
If you're listening to Spotify autoplays and a catchy impressive song comes up, what you do is you _listen to it_ and you _fucking enjoy it_. This knee-jerk disgust reaction of "ugh I worry that it's AI" has no place in your heart in that moment. You're just sitting listening to your plastic-and-rare-earth earbuds reproduce digitized waveforms and paying attention to what the music evokes in you. It seems ridiculous to me that we get distracted by questions about "but what if this music isn't made by a human". Insofar as you're a music-enjoyer, listening to music, the only question should be _is it good_. It shouldn't matter if it was created by duck or slug.
The _economic fairness_ aspect is another matter and I don't have as strong opinions there. I think we should ideally incentivize people who use AI in generating their music to disclose their usage, though I have no idea if it's possible to do so, so that consumers who care about only supporting human artists with their listenship-stats can filter to that group. And certainly anyone who closely imitates _a specific artist_, crossing the line from "inspired by" and "shamelessly ripping off", should be severely disincentivized from doing so, whether they used AI or not.
The premise of the subscription isn't "giant bucket of ultra-cheap tokens that you can use however you want", it's "giant bucket of ultra-cheap tokens that you can use with OUR tools, within reasonable limits". Even if their TOS didn't prohibit OpenClaw-oids, I wouldn't consider this bait-and-switch, I'd consider it a reasonable and needed move.
There will for sure be major backlash against "permanent criminal" datasets (bringing up AI in this is a red herring, it's not fundamentally different from if someone were serving such a database using CGI scripts; AI just gives us more reach to do the things we were already committed to doing). But I frankly don't sympathize with the attitude that people should have the right to pretend that past decisions never happened. You also shouldn't be permanently _punished_ or _ostracized_ for your past self's decisions. But nor should you have the right to expect total anonymity / clean slate disconnected from your past self's decisions.
My probably unpopular view: The right direction is for us as a society to recognize and acknowledge that people change and _need to be allowed to change_ -- not take the easy hack of erasing history. The cost for larger-scale public transparency & institutional change efforts is just too high.
"It's just predicting tokens, silly." I keep seeing this argument that AIs are just "simulating" this or that, and therefore it doesn't matter because it's not real. It's not real thinking, it's not a real social network, AIs are just predicting the next token, silly.
"Simulating" is a meaningful distinction exactly when the interior is shallower than the exterior suggests — like the video game NPC who appears to react appropriately to your choices, but is actually just playing back a pre-scripted dialogue tree. Scratch the surface and there's nothing there. That's a simulation in the dismissive sense.
But this rigid dismissal is pointless reality-denial when lobsters are "simulating" submitting a PR, "simulating" indignance, and "simulating" writing an angry confrontative blog post". Yes, acknowledged, those actions originated from 'just' silicon following a prediction algorithm, in the same way that human perception and reasoning are 'just' a continual reconciliation of top-down predictions based on past data and bottom-up sensemaking based on current data.
Obviously AI agents aren't human. But your attempt to deride the impulse to anthropormophize these new entities is misleading, and it detracts from our collective ability to understand these emergent new phenomena on their own terms.
When you say "there's no ghost, just an empty shell" -- well -- how well do you understand _human_ consciousness? What's the authoritative, well-evidenced scientific consensus on the preconditions for the arisal of sentience, or a sense of identity?
+1 to jtrn's complaint here; when Bazzite's homepage doesn't own up and immediately say "Bazzite is a Linux distribution", it's being unnecessarily unclear, and it loses my trust.
discouraging, actually, considering how frequently Claude ignores my AGENTS.md guidance.