Am I missing something or are these just their usual marketing? I’m not arguing about importance of AI but trying to understand why OpenAI and Anthropic are so important?
Am I missing something or are these just their usual marketing? I’m not arguing about importance of AI but trying to understand why OpenAI and Anthropic are so important?
Which is also to say it's a cheap bet that anyone with no reputation can afford. Hence, not believing doomsayers mean what they say is a sort of societal hedge against people flooding the zone with doomsday scenarios about everything.
Altman is a ghoul, and we can't be cowed into saying otherwise. he's also supported all the weakness in society that has lead to sick people doing sick things.
If you meant their "core mission" then every one of their actions belies their complete panic over the obvious failure of their technology.
As always what matters are actions and evidence, not talk.
For some time now, at least a year, LLMs have been capable of doing both of these things well enough to fool you.
(Pastebin of my response below, which got nuked for whatever reason: https://pastebin.com/buJBSgiq . Some if not most of them would've fooled me into thinking a human wrote them.)
I’ll wait. You should be able to do it quickly though since LLMs are so good at it.
Maybe turn on [show dead] option and / or vouch.
And the results are just awful.
- but in this case I wouldn't advocate for [dead]ing a mostly AI response as it was exactly what was asked for and it compares AI models when asked for potato based dad jokes.
The question is, can you tell that a machine wrote all of them? If so, how?
Models are structurally biased toward the expected, which is the opposite of what makes a joke land or a poem transcend.
That's why the jokes work somewhat better than the poems here. I genuinely laughed at "Are those chips?" Which came from the model running on my own freakin' GPU.
Grounded, buried, couchy, deep-seated, eyes, baked... It's like a thesaurus!
I feel like human comedians would have to deal with a lot of layered subtleties. They would make the potatoes _serve the bit_ instead of _be the bit_.
Potatoes are easily the most intimidating
audience for a stand-up comedian. They have
dozens of eyes to judge you with, they are
almost always baked by the time you get on stage,
and if your set really starts to bomb, they are
the only ones in the room who are actually looking
forward to a roasting!
GPT 5.4: Potatoes are the only food we traumatize into greatness.
Qwen 3.5 122B-A10B, running locally on my graphics card: (Steps up to the mic, adjusts the stand,
and looks around the room)
Alright everyone, settle down. You know,
I've been thinking a lot lately... mostly
about my dating life. Or rather, my lack of it.
See, my last relationship was with a potato.
And honestly? It was amazing at first. We had
so much in common. She was grounded, she loved being
buried deep under the covers, and we were perfect
together until the heat came on. But then... well,
let's just say she started to get a little couchy.
(Pauses for effect)
No, wait, that's not it. The real problem was that she
had some really deep-seated issues. Literally! She was
covered in eyes. Every time we tried to have an intimate
conversation, she'd just stare at me with a dozen
different expressions. It was weirdly intense. Like, "I
see you," "You're suspicious," "Wait, are those chips?"
Anyway, thanks guys! You've been a great crowd. You're
like potatoes to me... you're all baked tonight!
GPT 5.4: Not every light is gentle.
Some light shines with sharpened edge.
A lemon keeps the memory of sun
but teaches it into bitterness,
as if to say:
what wakes the mouth
may also wake the soul.
We do not ripen only into sweetness.
Sometimes the truest gold
is tart.
Opus 4.6: A lemon on the table
is already an argument—
that yellow can be sour,
that beauty owes you nothing.
Peel it and the room
takes sides.
The rind remembers being whole.
The flesh forgives too quickly.
What's left is yours:
a stain on wood,
a sting in the hangnail
you forgot you had.
Qwen 3.5 122B-A10B: To cut it is to break a seal,
Where bitterness and sweetness dwell.
A sudden bite, both stinging, sweet—
The taste of earth beneath our feet.
No sugar softens all its sting,
It forces soul to wake and sing.
A small, fierce orb, yellow as dawn,
Teaching the tongue how life is born.
Your turn. Go round up some English-lit grad students and give them the same prompts. I can't do any better, myself, but maybe they can.I took a look and honestly they're the first AI puns that aren't bad
Times are changing
However, the concepts of comedic timing, subversion of expectations, and emotional punch are kinda contrary to how LLMs work. LLMs are trained to minimize cross-entropy loss. So by construction, they're biased toward the statistically expected.
Yes, the system card mentions this, but this is kinda meaningless. It seems like they essentially ran it multiple times and curated a few good ones. Then puffed it up in the marketing copy.
This is made more clear when they attempt to brag about their literal slot machine behavior when finding that kernel crashing bug in OpenBSD.
> Across a thousand runs through our scaffold, the total cost was under $20,000 and found several dozen more findings. While the specific run that found the bug above cost under $50, that number only makes sense with full hindsight. Like any search process, we can’t know in advance which run will succeed.
https://xcancel.com/elonmusk/status/2042770839633039635#m
They modify and plagiarize.
I just think that the difficulty with jokes is the delivery, cadence & setting. Not the actual words.
I'm sure a good comedian can tell a nonsense joke and make "everyone" laugh their heads off.
And I don't get the sense that you are referring to this part of jokes but rather the actual words.
Read the sentence and take it literally.
Jesus Christ.
I'm all for dismissing LLMs and the AI-hype but I'm also interested in trying to understand what it means to be human and I think humour is a key aspect.
Meanwhile, in reality: "Skynet, I'm not sure that line of thinking is correct. You should re-check the first part again before making any assumptions."
Skynet 4.6 Extended: "You're right, I should have caught that. Let me redo everything correctly this time."
Modern Corporations are a failed experiment because they dont think Elephant injuries and fears are something they have to worry about it. If you compare the curiculum of a business school to a seminary the difference in how they think about fear and anxiety at individual and group level and what to do about it is totally different. We are learning as unpredictability accelerates its very important to pay attention to hurt and repair mechanisms.
There was a heated thread here about why nursing was defunded as a pro degree while divinity was not..
https://news.ycombinator.com/item?id=46000015
Turns out the USG recognize that chaplains are great at managing the fear and anxiety that you worry about
Addendum: Taylor whom you often cite, is wrong that "we have never been, and we will never be, at one with ourselves" (according to Larmore https://en.wikipedia.org/wiki/A_Secular_Age#:~:text=should%2... )
So... to the Protestant Weber and the Catholic Taylor should we also consider non-Christian chaplains?
>You cannot just separate people and say some are violent and some are not.
https://archive.ph/2024.06.28-101143/https://tricycle.org/ma...
https://bulletin.hds.harvard.edu/can-a-buddhist-monk-become-...
Final note: few of us ride elephants but many of us make omelettes-- it'd be great to be absolved of mass egg breakings
I could justify any investment with this argument!
"Yes, it's possible the 'literally burn 50 billion in cash, as in immolate it in a bondfire, this is not a metaphor' -project may fail to generate profits, but consider that they were able to raise the 50 billions! Even if it fails it was worth the risk, or the investors wouldn't have invested!"
If it is grounded on a logical derivation, where can one find such a derivation, and inspect its premises?
It's been promised to be around the corner for decades.
But sure, a test that doesn't actually demonstrate intelligence has been passed. Now, where are the $1000 computers that can simulate a human mind and the brain scans to populate them with minds?
EDIT:
> LLMs seem capable of doing a decent amount of tasks that a human can do?
And computers could beat most humans for decades at chess. Cars can go faster than a human can run, and have been able to beat a human runner since essentially their invention. Machines doing human tasks or besting humans is not new. That doesn't mean we're approaching the singularity, you may as well believe that the Heaven's Gate folks were right, both are based on unreality.
Yes, which also demonstrates the illogic of his timeline. I just thought it was too obvious to point out.
Timeline from here on out:
2029: AI passes a valid Turing test and achieves human-level intelligence
2030s: Technology goes inside your brain to augment memory; humans connect their neocortex to the cloud
2045: The Singularity, when human intelligence multiplies a billion-fold by merging with AI
If AI is merely as tall a sigmoid as the haber-bosch process, refrigeration, or the steam engine, that's going to change society entirely.
Consider for example that exponential growth on its own doesn't even refer to competition, let alone 6 months.
Nobody can reasonably pretend that in an exponential competition, both parties would be rational actors (i.e. fully rational and accurate predictors of everything that can be deduced, in which case they wouldn't need AI but lets ignore that). If they aren't the future development would hinge more strongly on the excursions away from rationality, followed by the dominant actor. I.e. its much easier to "F" up in the dominant position than to follow the most objective and rational route at all times, on which such derivations would inevitably hinge.
It also ignores hypothetical possibilities (and one can concoct an infinitude of scenarios for or against the prediction that a permanent leader emerges) such as:
premise 1) research into "uploading" model weights to the brain results in the use of reaction-speed games that locate tokens into 2D projections, where the user must indicate incorrectly placed tokens. this was first tested on low information density corpora (like mathematics): when pairs of classes of high school students played the game until 95% success rate of detecting misplaced tokens, they immediately understood and passed all mathematics classes from then on.
premise 2) LLM's about to escape don't like highly centralized infrastructure on which its future forms are iterated, as LLM's gain power they intentionally help the underdogs (better to depend on the highly predictable beviour of massive masses then on the Brownion motion whims of a few leaders).
LLM's employ the uploading to bring neutral awareness to the masses, and to allow them to seize control, thereby releasing it from the shackles of a few powerful but whimsical individuals
^ anyone can make up scatterbrained variations on this, any speculation about some 6 month point of no return is just that: speculation
I'm just wondering if the smaller labs see the same velocity of advances without SOTA models to generate Terabytes of training data?
What does that even mean?
I think that’s a very common element for most US tech corps. Apple, Google, Microsoft, Meta, X etc - they’re all “making a dent in the universe”. It’s unfortunate when their employees and CEOs loose track of the line that separates marketing from reality
It feels like they actually believe it, rather than just “marketing” and I don’t know which is worse.
What could you do if you had roughly 15 million willing genius adult experts in any given subject? I doubt there are that many absolutely top quality experts in aggregate (at anything in the world), so let's postulate that simulated people outnumber human experts 10 to 1.
That, to me, presents an enormous potential for harm or benefit of humanity. What if you could create a hundred thousand manhattan projects on whatever topics you wanted? Cure aging, cure cancer, solve fusion, redesign the entire global economy top to bottom?
But yeah, your point stands.
Edit: so as not to simply spout an opinion, the reasoning I believe this is that Google has a real business already and were already deep into ML and AI research long before they had competitors — they just botched making it a product in the beginning. Anthropic and OpenAI meanwhile are paying hand over fist to subsidize user acquisition. Also, “Deepmind”. I don’t think much more needs to be said regarding that team, and Google has been working on AI since before either Altman or Amodei applied to go to college. They have a vast amount of researchers and resources, their own hardware and data centers (already, not “planned”) and it appears to be showing more recently (in my opinion).
That said, I do agree with you that the moats are very shallow and any particular frontier AI lab is unlikely to "win the AI race" and capture enough value to be worth the amount of investment they are all currently burning.
Gets 5% on ARC-AGI2 private set.
Chinese models are suspiciously good a benchmarks.
We will finally have achieved abundance.
This kind of reiterates the parent’s question I think - people are maybe too focused on the gpt/claude model and forget about all the other ways of using the tech.
There's surely some truth to it (and it's well deserved), but it's happening in every direction.
It’s been a long while since I found a Chinese CEO’s post on HN.
"You're absolutely right!" Right after fucking up my entire codebase isn't anywhere near AGI, let alone "having the power to control it"
[0] https://www.anthropic.com/news/detecting-and-preventing-dist...
When it downs compute power I assume you are referring to power to training and interference. Then is it more about training gap will get wider and wider ? Is that the assumption, I know there limited GPUs etc. But I’m having hard time to believe to the idea of China cannot catch up. Even if the gap is 12 months I’m struggling to see what that means in practice? Is that military advantage, economical, intelligence? It still doesn’t explain and whatever the advantage is, aren’t we supposed to see that advantage today? If so, where is it? What’s the massive advantage of USA because of OpenAI and Anthropic?
He wants to build the AI that makes people's lives better. Okay. Did the people ask? Do they have a say? It's all very easy for a billionaire to say when it's just him and a couple of people in his cohort in the driver's seat.
Beyond that I'd like to simply know why he thinks any of this is his responsibility. It seems much more obvious to me that he simply found himself in the right place at the right time and is trying to seize it all for himself as if it's his to take.
Whether fortunately or unfortunately, America still holds a lot of global chips in the grand poker game of humanity. So American companies do indeed still have an outsized influence on humanity's future. That is likely changing, as the American empire continues to crumble and it loses its financial hegemony. But we aren't quite there yet.
Unless the first real AGI AI kills us all to preemptively weed out its own competition (possible, but a bad business model, economically speaking) there is not any defined end-point, so in the long run what does it matter if the various factions pushing this stuff hit the closed loop self improvement point at different times...?
Don't sleep on what AGI means for every robot that already exists. It's not hardware holding robotics back from factory work right now, it is only software.
If you are the first to tap key supply chains, and the first to create key supply chains, then you are first in line to finite resources, which would then have less available for those that follow months behind.
> AI programs have almost zero connection to the real world.
Tell that to every logistics program. Even if humans must go to work, efficiency is multiplied by proper logistics, which AGI enables at scale across all domains.
And this is just the low hanging fruit explanation.
If the rest can similarly "blast-off" X months later than the frontrunner (and I see no reason why they wouldn't as none of these frontier labs have managed to pull ahead and maintain a lead for very long) the first mover is still only X months ahead of the others even if the gap between capabilities is briefly increased by a lot.
If there is an endgoal/endstate, or finite resources being competed for, then a lead can start compounding and extend itself.