OpenAI builds first chip with Broadcom and TSMC, scales back foundry ambition
reuters.com
reuters.com
Also, did you know that Google's TPU efforts previously also relied on help from Broadcom? https://www.theregister.com/AMP/2023/09/22/google_broadcom_t... It's absolutely unsurprising that former Google employees would bring this vendor relationship into their new company, given Broadcom did a good job helping Google with TPUs.
BRCM is a leader in SerDes along with Marvell and other companies. There is so many IPs from Synopsys and Cadence that in turn many of these chip companies use. Not sure what is hoopla here. Nobody is going to completely vertically integrate all these IPs.
"while adding AMD (AMD.O), opens new tab chips alongside Nvidia (NVDA.O), opens new tab chips to meet its surging infrastructure demands"
Jensen has been saying that demand is "insane" and we're hearing rumors of low yields. This equates to supply issues in the coming months/years. No fortune 500 puts all their eggs into one basket. Diversifying away from a single source for all of AI hardware and software, is a smart thing to do.I wonder how this squares with the exclusivity contract with Microsoft. Even the OpenAI/Oracle deal requires Oracle to run Azure stack on their datacenter so MSFT can mediate the relationship. The AMD chips mentioned are also purchased by MSFT.
I wonder if this really means that OpenAI is accepting the risk/capital expense while providing a variety of hardware to Microsoft, or if there are other terms at play.
https://www.amd.com/en/newsroom/press-releases/2024-5-21-amd...
Not sure of the pricing, but the performance of their chip is pretty outdated at this point. MI300x is way faster and more memory.
That framing massively undersells how insane Sams ambitions were there, he was floating the idea of somehow raising seven trillion dollars to build thirty six fabs dedicated to making AI silicon. The TSMC execs reported more or less laughed in his face when he brought it up it them.
Yeah, that's um, wild.
You don't just acquire $7T.
The ENTIRE US domestic Net Investment isn't even $1T: https://fred.stlouisfed.org/series/W790RC1Q027SBEA
Gross Fixed Capital Formation (Net Investment + Deprecation) isn't much more: https://fred.stlouisfed.org/series/NFIRSAXDCUSQ
Google, Apple, and Microsoft together don't even spend $100B on CapEx per year. And they're worth almost $10T put together.
Asking for $7T when you're a $100B company is so ridiculous it's beyond belief.
And even them it would have to be split during a decade or two. And even then.
What's the military budget of USA?
Most of it is spent on personnel and operation.
You'd want to know only what the procurement is - which is ~$146B: https://en.wikipedia.org/wiki/Military_budget_of_the_United_...
And most of that is part of Gross Fixed Capital Formation already...
It seems totally possible that enough interest groups don’t want Ukraine to actually win, so they block anything more than the minimum viable amount. So it’s not much of a barometer at all for anything…
There's probably some small-ish 'dark budget'. But think about it: the official budget is in the order of single digit percentage points of US GDP. That represents and enormous amount of resources. Any appreciable fraction of that in a 'dark budget' would also leave traces all over the place.
The world is vastly more complex than say the world of 1945, so almost certainly there are now likely several dozens to several hundreds of ‘epicycles’ that are in fact true and necessary for full understanding.
Edit: Now the pentagon doesn’t contain the entire complexity of mankind within it… but it’s probably not that far off.
What bearing does what you mention have on the question on whether things got more or less complicated?
> Also, when expected outcomes don't align with perceived results, it's usually best but not very human to doubt yourself.
Eg this seems like something that would apply equally well at any time?
> Whoever has Shor/Grovers going first, has a advantage equitable to winning the AGI race... That too can win wars.
We have plenty of quantum resistant options, and people are already switching over.
And the story of the Enigma and its encryption shows that cryptographic and its breakage was already important back in the day.
> Nothing is simple and perceived simpleness is on the clock, for everyone.
How is this different now than before, though?
In monteary terms, this is barely more expensive than maintainig enough of a weapons stockpile during peacetime to counter Russia. And at the end of the conflict, anti-war sentiment within Russia might reach Germany post WW2 levels.
By comparison, more overt support to Ukraine, enough for Ukraine to at least humiliate Russia and throw them all the way out of the country, would only serve to help spread the anti-West hate within Russia, and generate support for even more ambitious rearmament in the years that would follow.
Pretty much like Germany after WW1.
> Accounting for non DoD military-related expenditure gives a total budget in excess of $1.4 trillion.[91]
https://responsiblestatecraft.org/pentagon-audit-2666415734/
Somewhere between $900B and $1T this year
I don't have papers but I think calculators, which are amazing and necessary, impact people's mental calculation capabilities. I know I used to he better at multiplication when i didn't have calculators all around me.
IMO, AI will have similar impact. Great advance overall but will nerf some areas of our mental capacity.
Like astronauts losing muscle mass faster in zero gravity.
AI impact on our brain, just like the calculator, might be a reasonable price to pay for the advancement it gives.
My manual calculation skills were at their peak in seventh grade in school, when we were allowed to use a calculator for the first time (and thus the exercises got harder) but simultaneously I usually forgot to bring mine.
But yeah, I don't think I could think better in general, when I did arithmetic better.
I thought it was really damning how Apple's recent ads involved a celebrity using it to respond to their agent about a script proposal they didn't actually read.
It just turns out that AGI (artificial general intelligence, which ChatGPT objectively is), is not a singularity nor even a tipping point towards causing one.
Put this another way, suppose we create a computer program that is only as smart as a bottom 10% human (I'm not saying they have.) You can't reasonably say that is smarter than humans generally. But are you comfortable saying that bottom 10% humans lack any general intelligence at all? General intelligence doesn't mean extreme intelligence, and so Artificial General Intelligence doesn't either. You might say that the term is more than the sum of the parts, which is fair, but I still dispute that superhuman abilities was ever part of the definition of AGI. It was just a (thusfar) failed prediction about AGI.
Now you could find a software that is smarter than some percentage of humanity in a few tasks. Is that software AGI? Is AlphaGo AGI? Is the Google Deep mind AI gamer AGI?
What you are getting confused about is that there is a segment of people that have purposefully confused / conflated ASI (artificial SUPER intelligence) with AGI, due to their own beliefs about the hard-takeoff potential of AGI. This notably includes the founders of OpenAI and Anthropic, and early backers of OpenAI prior to the release of ChatGPT. Through their success they've set the narrative about AGI, but they are doing so by co-opting and redefining a term that has a quarter century of history behind it.
The core mistake started with Bostrom, although to his credit he is very careful to distinguish ASI from AGI. But he argued that once you had AGI (the ability to apply intelligence to solving any problem) you would rapidly get ASI through a process of rapid iterative design. Yudkowsky took those ideas a step further in his FOOM debate with Robin Hanson in 2008, in which he argued that the AGI -> ASI transition would be short, measured in months, weeks, days, or even mere hours. In 2022, six months prior to the release of ChatGPT, Yudkowsky has a public meltdown in which he asserted there is no time left and we're all going to die, after having been privately shown an early version of ChatGPT.
It's almost two years since the release of ChatGPT, which is absolutely AGI by the definition of the people who coined the term and have run the AGI Conference series for ~25 years, or by the definition of Bostrom or Yudkowsky for that matter. ChatGP is general intelligence, and it is artificial. Two years of AGI, yet we are still here and there is no ASI in sight. Yudkowsky was wrong.
Yet OpenAI, which was founded by a bunch of Yudkowsky acolytes, is still out there saying that "AGI" will bring about the singularity, because that was the core assumption underlying their founding and the reason for the investments they've received. They can get away with this without changing their message because they've subtly redefined "AGI" to mean ASI, and are hoping you don't notice the slight of hand.
I'm sorry, but that's absolutely delusional.
I think for the people doing it, the expected endpoint is more about raising the price of shares they hold until they've been sold.
I wonder what they meant
LLMs have been shown to be capable of embedding logic and reasoning. There is nothing preventing a sequence generator from learning stochastic reasoning skills. Modern LLMs don't even just output language tokens. The latent information stored in the hidden layers is far more abstract than just language.
"We hypothesize that this decline is due to the fact that current LLMs are not capable of genuine logical reasoning; instead, they attempt to replicate the reasoning steps observed in their training data."
Here is an article where a model's learned ability to carry out modular addition is recovered. https://arxiv.org/abs/2301.05217
As you can see, the jury's still out.
Though really, all embedded models of reality are approximate by nature, and based on stochastic empirical data. The real tradeoff with probability-based neural networks is whether a rigid algorithmic approach or a more flexible, but stochastic, approach solves the problem at hand.
Then a week later it'll come up with a version of itself that's twice as smart. A week later that version will come up with a version that's twice as smart again, making it 4x as smart as a human. A week after that, it'll come up with a version that's twice as smart again, 8x as smart as a human. And so on. A year later it'll be a trillion times smarter than a human.
The AI will then either solve all our problems, invent fusion power and faster-than-light travel, and usher in a Star Trek style post-scarcity world of infinite resources and luxury, or it'll take over and create a Terminator style oppressive dystopia where the few surviving humans survive on rats we roast over a barrel in our underground bunker.
Elaborate on this idea please. Are you saying that human languages are self-replicating abstract entities?
The Singularity is the rapture of the nerds, and is about as likely to happen as the second coming of Jesus.
We see this in AI talk all the time "Dude is really good at coding and math, he's smart. We should totally listen to him when he makes things up like a 12 year old who just watched Terminator."
In the case of von Neumann, I mention him because he is the one who introduced the concept to the public. It is something he spent a great deal of time thinking about. He spent his life working around technology and had some incredible insights. Many consider him the smartest person to ever live, though in my opinion there are plenty contenders. He made key contributions to the development of computers, nuclear tech, quantum mechanics, mathematics and more. [0]
So I believe all of this truly does lead credence to the idea, at least enough to warrant sufficient research before dismissal. It's not his authority I appeal to, it's his experience. He didn't obsess over this idea for no reason, and his well-documented achievements aren't comparable to random flash-in-the-pan tech executives.
The ball is in your court to prove why the singularity is not real, because many experienced, decorated technologists disagree and have already laid out their arguments. If you can't prove that, then there's no argument for us to have in the first place.
If we don't grant that expertise on 1950s technology gives insight into 2040s technology, then there's little reason to consider his writing as something other than a cool story.
That's not my intention. It's more that, he simply looked at the historical progress of technology, recognized an acceleration curve, and pondered about the end game.
My argument is that it warrants consideration and not shallow dismissal. Plenty of opponents have criticized the concept, and some decent arguments revolve around rising complexity slowing down technological advancement. That is possible. But we can't just dismiss this concept as some fleeting story made up "like a 12 year old who just watched Terminator".
There are a lot of people out there who believe in technological progressivism and continued advancement of socially beneficial technologies. They just don't speak about a "singularity" beyond which predictive power is not possible, because the idea isn't worth spending time on as it isn't even self-consistent.
It's not our job to show why the Singularity won't happen. It's the nutjobs who believe in the Singularity who have the responsibility of showing why. In the 70 years in which people have been bitten by this idea, nobody has. I'll wait.
No, but shallow dismissal of a subject that many intelligent technologists have spent their life considering just comes off as arrogant.
Predicting technological timelines is a largely pointless exercise, but I wouldn't be hard pressed to imagine a future N decades or centuries from now in which, assuming we do not destroy ourselves, human affairs become overshadowed by machine intelligence or technological complexity, and an increasingly complex technosocial system becomes increasingly difficult to predict until some limit of understanding is crossed.
I read it on the internet myself with “sources saying” but I think it’s BS.
There’s a 68% chance that Altman met with someone who got their degree in the US (if only one executive attended, otherwise chances are higher).
And the others probably also have little difficulty communicating in English or calculating the cost of 36 fabs in any currency, since that’s their business.
Much more likely they talked about 3-6 fabs, some other person laughed about 36 fabs for podcast guy, someone else overheard that and posted on Weibo. That made it to NYT eventually.
>You don't just acquire $7T.
Perhaps the goal was only 1-2 Trillion but you should start high. I would have asked for 10 Trillion.....
My guess is that since they are so sure of AGI being achievable soon, that they think they can squeeze out any amount of money they want out of investors.
After all, AGI would be worth way more thab 7T right?
If you're applying for a job as a fry chef at a burger joint, asking for $25/hr is a good idea. Asking for $1000/hr will convince the shop owner that you are a crazy person who isn't serious about wanting the job.
It might be a $100B company on paper but that’s propped up by Microsoft investment. It’s hemoragging money still isn’t it?
https://www.theinformation.com/articles/openai-projections-i...
It took Google 15 years to build an ad empire.
And they still have trouble filling ad slots on YouTube.
You think ChatGPT is just going to "turn on" $100B in online ad spending overnight?
That's not how it works.
Personally, if I could buy OpenAI stock at 100B market cap right now, I would back up the truck.
Google can easily monetize a search for "dentist in $city" where your intent is to spend money, OpenAI would have to monetize "sum up this email" where money is not involved at all.
sure, but Google makes _way_ more money than X or Reddit or Pornhub, not all eye gazes are the same.
https://www.fierceelectronics.com/electronics/openai-ceo-alt...
> He added the number is not known, adding that the $7 trillion came from a report with an anonymous source. “You can always find some anonymous source to say something, I guess,” he added.
Lobbyists?
I'm not personally knowledgeable about this or anything, so look into it yourself, but just picking a random really big one, United Airlines, it looks like they consistently beat the S&P 500 for most of the past 20 years, until Covid, and the past four years since then have been pretty bad.
What they don't seem to be able to weather is significant changes in the structure of pension obligations, and dramatic fatal crashes.
There is also an impending shift into electric short-haul flights that is going to cause some changes to the business model, with whoever can't adapt exiting those markets.
That was why you had such rapid consolidation, as America West became US Airways became American, United bought Continental, and Delta bought Northwest, to remove so much competition that they could be profitable again.
I read that thing about the integral of all profits back to the Wright Brothers back around in 2010, and I can't find a good source on aviation profitability in the 1970's right now.
https://www.buzzfeednews.com/article/richardnieva/worldcoin-...
https://www.technologyreview.com/2022/04/06/1048981/worldcoi...
Especially not one whose business plan is “let’s invent fairytale technology then ask it how to make money”.
https://x.com/geepytee/status/1778859803735113950
If all of that sounds good to you, I know of a bridge you may be interested in purchasing.
The startup playbook lie ur ass off and make sh!t up. Musk still does it to this day with the recent demo of his Optimus robots implying they were all AI driven.
Some with deadly consequences like Uber's attempt at self driving and what was that other recent self driving company that ran over / mangled a pedestrian?
Tbh when they had the real person on the suit I wouldn’t have approved that it would give the impression I’m ok with pretending to an extent.
I understand having a half working prototype iPhone on stage but literally pretending the robots are real just felt dishonest and a bad idea.
But hey I’m not a Vc
If you spoke with the robot and believed it was AI controlled you should feel like a moron, considering the robot explicitly stated that it was remote controlled in that demo. https://x.com/zhen9436/status/1844773471240294651/
1) "that" demo? I never mentioned "a" demo. That one clip, which you found, the robot answers a specific question to one user.
There were hundreds of interactions that night, and in many that I saw atleast one person clearly didn't realise they were.
2) Regardless of whether it was their fault for being morons, my point is they FELT like morons, which TO ME is not good publicity.
I like Elon and I think the demos were incredibly, I am saying it weirded me out they were trying to pretend the robots were sentient, which I am sure they are not far away from. I just think it was short sighted. But maybe I give people too much credit.
I would 100% have not told anyone to pretend shit. It wasnt even required. It just made me feel uneasy about the products.
If he had said $250 billion and six fabs, it would have been a lot to ask but people wouldn't think he was ignorant or irrational for saying it. Big tech for example has that kind of money to throw around spread out across a decade if the investment is a truly great opportunity.
Could it be possible that OpenAI's new autocomplete will be as transformative to the global economy as WeWork's short term office rentals?
When looking at things like mechanization and the productivity increases from around 1880 to now, you took an economy from around 10 billion a year to 14 trillion. This involved mechanization and digitization. We live in a world that someone from 1880 really couldn't imagine.
What I don't know how to answer (or at least search properly) is how much investment this took over that 150 year period. I'm going to assume it's vastly more than $7 trillion. If $7 trillion in investment and manufacturing allowed us to produce human level+ AI then the economic benefits of this would be in the tend to hundreds of trillions.
Now, this isn't saying it would be good for me and you personally, but the capability to apply intelligence to problems and provide solutions would grow very rapidly and dramatically change the world we live in now.
Sure, but right now that "if" is trying to do 7 trillion dollars of unsubstantiated heavy lifting. I might be able to start creating Iron Man-style arc reactors in a cave with a box of scraps and all I ask is 1 trillion, so you all should invest given how much money unlimited free energy is worth.
The words "just" and "surely" are doing so much heavy lifting here that I'm worried about labor-laws. :p
If scaling up was "just" that easy and the next step would "surely" begin in the exact same architecture... Well, you've gotta ask why several billion general-intelligence units made from the finest possible nanotechnology haven't already merged into a super-intelligence. Or whether it happened, and we just didn't notice.
Reminds me of:
Musk flew there with Cantrell, prepared to purchase three ICBMs for $21 million. But to Musk's disappointment, the Russians now claimed that they wanted $21 million for each rocket, and then taunted the future SpaceX founder. As Cantrell recounted to Esquire: “They said, 'Oh, little boy, you don't have the money?”
Source?
when i read that news back then, i read it as "billions" and thought "that sounds very ambitious, hardly he'll get that money". Man, i'm 3 orders of magnitude behind the current state of the art in tech, and need to start thinking in trillions (while it was only recently when a billion dollars was large money).
It's a bold strategy, Cotton. Let's see if it pays off for 'em.
0 - https://www.wheresyoured.at/subprimeai/>>OpenAI launched o1 — codenamed Strawberry — on Thursday night, with all the excitement of a dentist’s appointment. Across a series of tweets, Sam Altman described o1 as OpenAi’s “most capable and aligned models yet.” Though he conceded that o1 was “still flawed, still limited, and it still seems more impressive on first use than it does after you spend more time with it,” he promised it would deliver more accurate results when performing the kinds of activities where there is a definitive right answer — like coding, math problems, or answering science questions.
it seems to me like they nerfed the regular free ai model to give me imperfect replies on my benchmark consistently while the o1 model nails it consistently. before the free model would, depending on the day, give me more accurate replies to my benchmark .
thus widening the divide to make the o1 model look more useful. openai can do whatever they want and thats fine but all it seems to me they did was make 01 look good by comparison by nerfing the free product.
My main takeaway from the article was more about the unsustainable business model due to a massive cash burn rate. To wit:
OpenAI’s models and products — and we'll get into their
utility in a bit — are deeply unprofitable to operate, with
the Information reporting that OpenAI is paying Microsoft an
estimated $4 billion in 2024 to power ChatGPT and its
underlying models, and that's with Microsoft giving it a
discounted $1.30-per-GPU-an-hour cost, as opposed to the
regular $3.40 to $4 that other customers pay. This means
that OpenAI would likely be burning more like $6 billion a
year on server costs if it wasn’t so deeply wedded to
Microsoft — and that's before you get into costs like
staffing ($1.5 billion a year), and, as I've discussed,
training costs that are currently $3 billion for the year
and will almost certainly increase.
Whether or not the "free model" is better or worse than the "o1 model" is moot IMHO.You can throw more money at it to make it go faster, but it also might fuck it up and take longer.
Which is the story of AMD for the last ~15 years.
getting support so we could develop apps for their graphics card was dispiriting. at the time they were faster and very much cheaper than nvidia (but no cuda) But every time we found an issue, the poor devs who were contracted to fix it were left struggling.
Its the same where I am now. We were trying to qualify some motherboard/TPM/other issue, but they didn't have enough bandwidth to help us in time.
Its better now, we do have some AMD stuff in the fleet.
At least from someone who is on the edge of the business and seeing what is going on. It is a lot better now, especially at the enterprise level. They've been hiring and buying companies too.
It won't be fixed over night, but it is definitely a renewed focus.
Once they get it right, they will increase the gap with other LLM providers.
Apple didn't build their M1 chip from scratch. By the time the M1 chip was released they had already been building chips for the iPhone and iPad for several years!
Are any of the western AI/matmul chips NOT being fabbed by TSMC? Amazon or Meta's perhaps? We've got NVIDIA, AMD, Google (TPU - also designed by Broadcom), and now OpenAI all using them. Are Samsung not competitive?
I wonder to what extent total chip/chipset volume is limited by TSMC capacity as opposed to memory or packaging?
If they had proprietary fabs and chips that were magically superior, and they could build them for cheap rather than paying the team green tax, they'd have a huge advantage in inference cost.
Perhaps they've realised just starting with a proprietary chip is enough?
Ditching it might mean ditching the tooling that makes it so effective.
I would imagine some productive cultural impact that’s wide spread, sticky, high retention not like ‘oh cool 3d tv or vr’ kinda thing?
Given how much ‘money’ has been burnt through the GPU heatsinks.
Apart from people in education, a whole SEO industry has grown around LLMs and they love the newly found powers to produce spam at scale.
In business / government space people are cautiously introducing LLMs to improve productivity in the comms area where slight hallucinations or robotic corporate language are acceptable, because the final text is still produced by a human. On the customer support / sales side companies are investing money in "AI" versions of the good old phone menu trees that inevitably make you scream "I want to speak to a human" before you get connected to a human being.
Elsewhere adoption is slow, because we expect deterministic results and these are just not to be expected. I write software that interacts with LLMs via APIs and have to implement defensive coding techniques I haven't used for over a decade, because the output may switch from the requested structure at random. This makes use of structured output or function calling problematic.
Once all outlook users have buttons to summarize, extract into, improve the email, then you have hundreds of millions using LLMs everyday.
I also used to be in disbelief, but i'm seeing more and more unexpected people use LLMs , and they even figure out how to prompt them right.
Particularly o1-preview is really good at writing broad strokes mini review style articles that discuss some question you're interested in.
It only works in cases where there's a lot of text out there already that discusses the topic and that needs to be sort of remixed to answer your question specifically.
The chat format is much more efficient for this than a web search or reading a few related Wikipedia articles.
It's truly glass half full vs. glass half empty. I'm super impressed and excited about how useful this is, and yet it clearly isn't AGI.
It seems pretty clear to me that this is disruptive to most knowledge worker's workflows, even without getting into the whole discussion of how close are we to AGI and does this thing really understand anything.
It's basically Clippy done right ;)
A highly advanced content retrieval system (but lacking in accuracy, so ok for situations where that may be ok).
I wouldn't say it's revolutionary -- certainly not as revolutionary as search, but it's a major step forward in being able to retrieve information. Ultimately will allow you to interact with computers the way you see in the old Star Trek movies, where you're speaking in plain English and receiving information back in the same (providing the systems which are the source of the needed information are interconnected and fed into the training set).
If not, will everyone else have to build their own chips to compete?
So they planned to transform everything into paperclip factories?
... Wait, what, _really_? _Why_? I mean, it'd probably take about a decade to get going.
Splunk+Datadog and AWS+Nutanix+Cohesity respectively ate much of CA and VMWare's marketshare.
I don't think that's completely wrong, but it's a big company and I'm sure there are some better areas of the company than others.
General consensus on HN is generally wrong.
I think everyone here would agree with that.
;)
HN doesn't understand business in general. I miss the good old days on HN when you actually saw execs or actual SMEs shooting the shit.
Now it's just Reddit and LessWrong refugees based on account creation date.
It's a sign that we're moving on to greener pastures.
Sucks for the new generation that think everything needs an app to work.
The internet was better at feeling less corporate a decade ago.
Having been on both sides, I’m continually shocked that stuff even works.
Avago Technologies (owned by corporate raiders) bought Broadcom, then took its name for itself
its ticker is still AVAG
So both are correct. Broadcom is a legit important hw company, and it is a PE holding company that buys and loots tech companies (like VMware)
>TSMC execs allegedly dismissed Sam Altman as ‘podcasting bro’ — OpenAI CEO made absurd requests for 36 fabs for $7 trillion