HNHacker News
TopNewBestAskShowJobs

claytonjy

1,497 karma · joined July 22, 2014

startup ml engineer
submissionscomments
claytonjy··on Was modern art a CIA psy-op? (2020)
Betteridge’s law prevails. A fun read, but not as fun as the title suggests.

tl;dr Modern art existed before the CIA and even CIA predecessors did not originate it, but they did push it abroad to show that America is great because we put up with modern art made by socialists depite most americans disliking both the art and the artists

claytonjy··on Coding expertise is going to collapse from AI reliance
They do, it’s called the AMA, American Medical Association. Very tight control of who can be a doctor, how many new doctors per year, etc. which ensures salaries stay high. They also fight against other roles like nurses doing too much doctor-like work, helping doctor pay while harming nursing pay.
claytonjy··on How Kubernetes Probes Work
The author re-implemented a bunch of k8s logic in a typescript library just for these animations: https://github.com/ngrok/webernetes
claytonjy··on Nvidia's Risky Business
you might be looking for SemiAnalysis? I only read the free portions of articles but they have various paid options, mostly targeting investors with information and tools.

https://semianalysis.com/

claytonjy··on Uber SubmitQueue: a high-performance speculative merge queue
I haven’t worked at a Big Tech monorepo place, but I hope/assume that in the service dependency case, you would touch both in one changeset and the deploy tooling understands the dependency tree and orders the deployments correctly. Without that i agree its the polyrepo approach just with all PRs in the same repo instead of multiple.

In the library scenario, I know tooling like bazel ensures that’s one changeset, not 3+. Tests run for both the changed library and all consumers of it in the same pass. You’re right that it might loop in others for review who weren’t expecting it, but i think that’s the same in a polyrepo approach.

claytonjy··on Uber SubmitQueue: a high-performance speculative merge queue
if service A calls service B, and service B adds a new endpoint, or a new optional argument, service A needs an update to take advantage of it

if a library is used by multiple services, and gets an important bug fix, each service using the library needs to update to get the fix

these are sequences of changes, not literally at the same time or requiring deployment coordination, but when this happens a lot people start asking about monorepos

claytonjy··on Making Postgres 300x faster for analytics: batching, operator fusion, and SIMD
does one not imply the other? if it can run faster in the same hardware, it should also run as fast on lower spec hardware
claytonjy··on Situational Awareness and the Impending Stock Market Volatility
the last step, #7, is that fully automated or is this humans calling humans? I imagine everything before then is quite automated, and are thus happening very quickly, so I'm curious if the last piece possible being manual has the potential to blow the whole thing up by being too slow.
claytonjy··on Why Software Factories Fail (or: harness engineering is not enough)
> There are 'points of view' that emerge during coding

I know it’s a bit cliche at this point, but this harkens to “programming as theory building”[0] which I agree is easy to lose out on when embracing agentic coding today.

[0]: https://gwern.net/doc/cs/algorithm/1985-naur.pdf

claytonjy··on Why Software Factories Fail (or: harness engineering is not enough)
thanks for pointing out the linear PR stuff, hadn’t seen that. Interesting that while a dozen other companies are trying to muscle in on the hosting/versioning side of github, rather fewer are working on the PR side.
claytonjy··on I Inspected My Take-Home Interview Project. It Was a Whole Operation
Isn’t this how most tech companies, if not most companies in general, operate? Ive only had one BYOD job in 15 years. Even if you’re a contractor this is common.
claytonjy··on Nobody knows what a used GPU cluster is worth
I think there’s a bit more nuance and complexity here.

I built and operated an application that used 1-2000 L4 GPUs in production for a couple of years. Long-term GCP reservations running in a GKE cluster. At that scale we had a few GPU failures per week, and once saw 3 in a day. Nearly every failure required an engineer to manually intervene to get rid of the bad node, and then file a ticket with GCP support as they requested.

NVIDIA GPUs throw an “XID” code when they fail, which can be seen from serial port logs for GCP compute nodes (not through k8s!). If you’re lucky, it fires right as your application starts to fail, but there’s often a delay of several minutes. Even when you get an XID, by default GKE only responds to one or a few of them. You can expand the list via configuration, but the reconciliation loop is so slow that might still take 10-15 minutes during which a pod is puking errors and someone might be getting paged.

They’re working on it, and we never saw a single XID after we migrated to H100s, so the situation is improving. I imagine other clouds are even worse, though iirc Azure was leading some effort to improve k8s node problem detector to include accelerator problems so maybe they have a better story.

Training workloads are rather more sensitive to a node failure given that many modern training runs (including SFT etc) need multiple nodes where topology matters, and they might not have another e.g. 8xH100 box in the right place when one fails.

claytonjy··on Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
How big is this market, self-hosting a model that requires 64 GPUs, H100 or better, with good interconnects between nodes?

I suspect the overlap of those that can afford it, and those that have the talent to manage it, is a fairly thin slice of the Venn diagram. Even the large corps are gonna be getting it from the inference vendors, or more likely Bedrock and friends.

claytonjy··on The Korean telecom giant at the center of Anthropic's Mythos controversy
I can also use the models anytime, and for a lot of time, until i’m anywhere close to salary for an engineer like that
claytonjy··on Stop Using Conventional Commits
release-please[0] allows you to do a manual version override in a commit, which would allow you to decrement the major version upon reverting a breaking change

I think that could be simplified, so the tool can tell that a commit is reverting a breaking change and thus the version should be decremented, but at least there's an escape hatch.

[0]: https://github.com/googleapis/release-please

claytonjy··on Sagrada Família Lego set
With a lot of the adult-oriented sets (ICONS and others), especially anything plant-based, they go out of their way to point out how they're using existing pieces in an unusual way. One example is the cherry blossoms on the bonsai tree are actually frogs, but cast in a pink color for the first time, called out in the instruction manual as you're building it.

I love it, knowing about these little details. Also fun to share with friends that inquire about the various LEGO on display in the house. This, and all the fancy mechanics (e.g. typewriter, nintendo), engender a ton of respect and awe for the designers.

claytonjy··on Bun's experimental Rust rewrite hits 99.8% test compatibility on Linux x64 glibc
Kind of the opposite, I was deep in the R world a decade ago and there was a huge trend of replacing Java dependencies with C/++ ones because the JVM was such a pain to manage. The community eagerly adopted the replacements about as soon as they existed.
claytonjy··on Today I've made the difficult decision to reduce the size of Coinbase by ~14%
I experienced a flavor of this, too. We had some outages, management said no more daytime deploys, so we had after-hours “deploy parties” whose scope and participant count increased weekly. The smarter managers said it was temporary, but couldn’t say how we’d move back towards continuous deployment. If anything went wrong in any service, you’d end up with a dozen or so folks on a zoom call for 3 hours. We did this once or twice a week.

Went on for about a year, worse each week, before i left.

claytonjy··on Write some software, give it away for free
I don’t know how big the market is, but seems pretty commercial-friendly to this old magic player. I have a big box of cards from a few decades ago I’ve held onto. I’ve thought about selling them, but it seems i either take them to a shop and get lowballed, or spend hours meticulously researching each card and then figuring out how to sell it for what it’s worth. taking a pile of photos and having the ID and valuation automated could go a long way! Hard to sell to individuals like me, but i would think a card marketplace would find it invaluable?
claytonjy··on Computer Use is 45x more expensive than structured APIs
What sorts of sites are you thinking of? To me, “most useful to a programmer” evokes docs and blogs and github issues and forum posts. I suppose some forums might be AI-resistant (login wall), but the others are trivially AI accessible.
claytonjy··on Today I've made the difficult decision to reduce the size of Coinbase by ~14%
Yeah this sounds pretty reasonable really, like instead of using a CMS directly they’re having Claude file PRs to make the same changes. As someone who likes static sites and change control, it actually sounds like an improvement.
claytonjy··on Today I've made the difficult decision to reduce the size of Coinbase by ~14%
It is, but it’s the only way for a company to succeed and scale over time. A pet approach works well in the early days, but you can’t become a VC-backed success without drastically reducing bus factors throughout the company.

That could be an incentive to keep companies small, but high-scale companies do have unique benefits to society.

claytonjy··on Today I've made the difficult decision to reduce the size of Coinbase by ~14%
My experience as well. It sounds nice at first, but since it’s tied to org flattening these “player-coaches” end up with 15-20 reports, which is way too many for even a pure manager.

I noticed it was especially bad for on-call and incident response; these managers get pulled in to all the incidents because of their status and supposed involvement, but are not particularly useful in those rooms, adding even more cooks to the already crowded kitchen.

claytonjy··on GitHub Stacked PRs
This is what i often do, but i have never been able to get many coworkers onboard. In my experience I’d say less than 5% of all software folk i’ve worked with are willing to do an interactive rebase; everyone else finds it too scary
claytonjy··on Launch HN: Chamber (YC W26) – An AI Teammate for GPU Infrastructure
> most teams we talk to can't even tell you how many GPUs are in use right now

how can this be? isn’t this a trivial metric to pull from any clouds monitoring service?

to get the good ones (H100+) you generally have to reserve them, a fixed cost you pay monthly and can’t pretend to not know

claytonjy··on Pandas 3.0
Even before LLMs, Data Science was being replaced by more specialization, IME.

Data Engineers took over the plumbing once they moved on from Scala and Spark. ML Engineers took over the modeling (and LLMs are now killing this job too, as it’s rare to need model training outside of big labs). Data analysts have to know SQL and python these days, and most DS are now just this, but with a nicer title and higher pay.

Once upon a time I thought DS would be much more about deeper statistics and causal inference, but those have proven to be rare, niche needs outside soft science academia.

claytonjy··on Google, Nvidia, and OpenAI
not even google thinks this will happen, given their insistence on only offering TPU access through their cloud
claytonjy··on Claude Code 2.0
plans now open in a separate file tab, and if you don’t accept it, it just…disappears so you can’t discuss it!
claytonjy··on Providing ChatGPT to the U.S. federal workforce
you can’t really buy H100s except in multiples of 8. If you want fewer, you must rent. Even then, hyperscalers tend to be a bit inflexible there; GCP only recently added support for smaller shapes, and they can’t yet be reserved, only on-demand or spot iirc.
claytonjy··on Fast
We have common words for those two flavors of “fast” already: latency and throughput. S3 has high latency (arguable!), but very very high throughput.
Page 1 of 23Next →