HNHacker News
TopNewBestAskShowJobs

dylan522p

195 karma · joined June 2, 2021

submissionscomments
dylan522p··on Amazon’s Cloud Crisis: How AWS Will Lose the Future of Computing
The article does answer the question, in detail. What you're reading is just the first half is freely available. The 2nd half is on the same web page and for subscribers.
dylan522p··on Google Apollo: The >$3B Game-Changer in Datacenter Networking
They use some of this too btw. Also wavelength level routing happens with breakout cables from ToR to Compute.
dylan522p··on Embracing Chaos: The Imperfect Art of Semiconductor Manufacturing, Lithography
OPC? It'd be great to talk about your area, because I bet most don't know what you do, and maybe I can make it an entertaining read.
dylan522p··on The Inference Cost of Search Disruption – Large Language Model Cost Analysis
No TPUs are getting better architecturally
dylan522p··on The Inference Cost of Search Disruption – Large Language Model Cost Analysis
Your math is completely wrong dude.

It's 2000ms per token not for the whole query.

Hardware utilization rates and MFU are not the same thing, you forgot the latter.

You are pretending its perfectly parelelized on 1 GPU too. I use 8x GPU box throughput.

dylan522p··on The Inference Cost of Search Disruption – Large Language Model Cost Analysis
> This article is based on wild speculations of how much things cost

Huh it literally uses real throughput figures

> doesn’t account for ways to make things cheaper over time

It does in the subscriber section and it says it does say that in the free section.

> We already have papers that suggest most big models are undertrained and smaller models can get the same accuracy.

Why assume 2023 model is the same as 2020 GPT-3 175B parameter

dylan522p··on The Inference Cost of Search Disruption – Large Language Model Cost Analysis
- Google qps is closer to 100k then 320k [1]

That number is wrong. I have a number from googler, not livestats which cannot have google internal data

- Not every query has to run on LLM. Probably only 10% would benefit from it

Agree, i have something different coming up that looks into this more, 10% may be too low. I know i used 100% which isn't right, and explicitly say that

- This means 10,000 queries per second, each needing 5 A100s to run, so 50,000 A100s are sufficient. Cost for that is $500MM, quadruple that to $2B with CPU/RAM/storage/network. That is peanuts for Google.

50k A100s networking ramp cost way more than $2B HW utilization rate

- Latency, not cost, is a bigger issue. This should be addressed soon by H100 and newer chips.

Thats discussed in the subscriber section. It's both, but yes latency is bigger issue. H100 helps but doesn't solve.

dylan522p··on How Nvidia’s CUDA Monopoly in Machine Learning Is Breaking
cutlass is faster actually, in most cases where cutlass supports same stuff.
dylan522p··on How Nvidia’s CUDA Monopoly in Machine Learning Is Breaking
Thanks! Ya, the archive link is useless, it only captures the free part, and I think I am very generous with what I keep in the free section on the ad-free website.
dylan522p··on China’s SMIC Is Shipping 7nm Foundry ASICs
Tsmc 7nm doesn't use EUV and is fantastic
dylan522p··on Apple M2 Die Shot and Architecture Analysis – Big Cost Increase and A15 Based IP
Author here. What do you mean die shot bleeding?
dylan522p··on Apple M2 Die Shot and Architecture Analysis – Big Cost Increase and A15 Based IP
M1 2020, M2 2022?
dylan522p··on Apple M2 Die Shot and Architecture Analysis – Big Cost Increase and A15 Based IP
It has noticed. Look at Apple architects, validation, layout, etc engineers moving to Nuvia + Rivos + Google + Amazon + Microsoft + Meta + Intel + Nvidia + AMD + Apple + Qualcomm.

It's there.

dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
It's at the end of every article lol
dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
> the unsubstantiated BS that Samsung's chip manufacturing is a disaster.

They lost Qualcomm, Nvidia, and Cisco for the next generation. Are you disputing this fact?

They did not fully ramp 1Z or 1 Alpha dram nodes. Are you disputing this fact?

dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
I have been sent C&D letters in the past by Arm, and even sued by others. I have resources.
dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
If you don't want to believe it. Go ahead. The major claims are true, Qualcomm moving away. Nvidia moving away. DRAM node ramps being pitiful. DRAM engineering efforts come directly from a source there.

The cultural issues being the cause of these issues is the substantial claim.

dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
MediaTek is killing it. They hired a lot of people from TSMC's process development kit teams which vastly improved their capabilities in integrating IP. They were the first to develop AV1 decode, their modems are pretty decent. They even have leading teams in wifi.
dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
Most folks blame the culture shift on Paul actually. He was the cause of it all. Brian being a horrible symptom of what he created, and Brian definitely propagated it far.
dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
If anything I am making the case that TSMC is reigning supreme but okay, believe your conspiracy
dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
Good thing the article even says Samsung is a well oiled machine in many areas outside of the ones with issues discussed in this article. Almost like large companies can have varrying success and culture.
dylan522p··on Samsung Electronics cultural issues causing disasters in foundry, LSI, DRAM
It's at the bottom of every article. I have to disclose given my relationship with various funds.

I am under no impression that my articles can move fucking Samsung's stock. That's hilarious you think I could profit off the market by writing this.

There's plenty of evidence though. Be an industry insider and you'd recognize it all.

I do know that my reports have moved smaller companies stocks by 20% in a single day, and have been verified true in the past. - https://semianalysis.com/short-report-nvidia-supplier-cut-ou...

If I thought I could move the stock, I'd make the position in the morning alongside my clients, and publish shortly after, like I did with the article I just linked.

dylan522p··on As Moore’s Law slows, Apple is forced to use cheaper chipsets in non-pro iPhones
My original article you are talking about explicitly stated heterogenous compute is the future. It talks about transistor budgets going to LLC, GPU, media, and ISP. Much of the CPU efficiency gain comes from that doubled LLC size.
dylan522p··on As Moore’s Law slows, Apple is forced to use cheaper chipsets in non-pro iPhones
Sophie's analysis contemplates a set unit volume and fixed costs related to design and verification and tape out. Mine is on pure cost/transistor based on wafer costs as a player like Apple has the units to spread fixed costs over.
dylan522p··on As Moore’s Law slows, Apple is forced to use cheaper chipsets in non-pro iPhones
Look at the in depth analysis I did of the die shot and SOC floorplan. Or even read the article I originally wrote I was right on the CPU architectural gains. Apple pushed clockspeed up slightly, but IPC was less than 5%.

https://semianalysis.substack.com/p/apple-a15-die-shot-and-a...

dylan522p··on As Moore’s Law slows, Apple is forced to use cheaper chipsets in non-pro iPhones
Why no on the BOM increase? That is what the switch to LP5 and A16 would be.
dylan522p··on As Moore’s Law slows, Apple is forced to use cheaper chipsets in non-pro iPhones
The IPC gains were what I specifically wrote about and those turned out to be true with <5% IPC gains on the Avalanche core CPU
dylan522p··on Arm China Has Gone Rogue
They are allowed to sell chips outside of china with Arm IP licensed from Arm China
dylan522p··on Arm China Has Gone Rogue
The last part of the article was not here last year. The last part and images are from an event they held recently. Notice the

"Before we get to the event they held and the significance of it, let’s do a recap."

He refuses to leave and he has the stamp. The 7-1 vote was even mentioned.

I would love if you could find those images from the event last year. You would need a time machine for that.

dylan522p··on Arm China Has Gone Rogue
Arm does not manufacture chips. They license IP for per chip or blanket fees. The model of many of these licenses is irrevocable. I've seen a couple different Arm licensing contracts and they're all very different so hard to make blanket statements. The Chinese entity has the right to license to all Chinese semiconductor firms who have the right to sell their chips.
Page 1 of 2Next →