HNHacker News
TopNewBestAskShowJobs

metalspot

360 karma · joined December 2, 2022

submissionscomments
metalspot··on Does Reddit have an astroturfing problem? What the data suggests
> this is no longer an indicator of a bot account

it is trivial now to create synthetic multi-platform personas that generate real looking post histories at unlimited scale. the baseline assumption should be that everything on the internet is fake until proven otherwise.

metalspot··on Coding is not solved
> Reading the code does not mean you understand the code

this is delusional hubris from people who are not real engineers. anyone who actually writes high assurance software or does low level performance optimization knows that the only thing that matters is what can be demonstrated in reproducible tests.

the full stack requires about a 1000 different specialties that each require about 10 years of experience to master. and that is just one computer. we are building distributed systems of millions of these computers, operating at global scale, across dozens or hundreds of legal jurisdictions, which takes the complexity of a single computer, and multiplies it many times over, across several other dimensions.

anyone who thinks that they can understand the systemic effects of changing 0.00001% of this system by reading the code is a dangerously naive fool.

metalspot··on Coding is not solved
You aren't coding though. The LLM is coding and you are specifying.

It is just a function. A next token prediction function. Garbage-in/Garbage-out is as true as ever.

Coding is solved. Specification and translating requirements into specifications is an unbounded domain and unsolvable by definition.

What LLMs change is that now you develop the specification through iterative implementation and testing.

You start with a weak specification which produces slop, then you progressively build the specification with different approaches to automated testing. The tests are the specification. Ultimately, what you ship is the tests. They are the only thing that proves the functionality of the program. An LLM can produce code that passes any test suite you give it. If the result is bad, the problem is not the code, the problem is the specification.

metalspot··on Coding is not solved
There seems to be a big disconnect here depending on level of experience.

This is one of the most senior engineers at Amazon saying that human code review is dead: https://x.com/MarcJBrooker/status/2101005954708021604

Nobody at this level is "vibe coding." They are using LLMs as a tool in a whole suite of tools that they have spent decades mastering, and LLMs happen to be the most powerful tool ever created. When you understand the existing tool suite, and can integrate this new super tool, the results are order-of-magnitude improvements in velocity, with much higher levels of quality and assurance.

Anybody talking about "code," like it is actually important, simply lacks the perspective to understand this.

We went from punch cards, to assembly, to C, to interpreted languages, to frameworks, to AI, and every cycle had the exact same debates.

A lot of it is Ego. Everyone thinks they are smarter than they actually are and that their work is uniquely valuable.

Computer programmers are monkeys who get paid to press buttons. We get paid because we know which buttons to push and in what order. It's a great gig. It's made me more money than I ever imagined possible, and I have fun doing it, but the flip side is that it is the most competitive industry on earth.

If you slow down, fall behind, and refuse to adapt, you will get eaten alive.

I have some sympathy for people, but an the end of the day, if you want to get paid better that 99% of people on earth, you are going to have to work for it. That is not an entitlement, and if you think it is, you will not make it.

metalspot··on Coding is not solved
The audacity of publishing self-promotional AI slop clickbait claiming that AI can't code and everyone who doesn't agree with your asinine assertions is incompetent is bold. Respect the hustle I guess.

But to anyone even vaguely thinking of taking this seriously, go look at what antirez, dhh, jared sumner, mark brooker, and many other real engineers who have ship real things are doing and saying.

Most of these people have spent their entire lives contributing to open source, and they have proved their skill shipping working software and scale for decades. They are really trying to help people by showing and telling them exactly how AI works and how to use it to make better software.

metalspot··on There is more to code review than (automatable) detection
> and is this a solved problem?

yes. it was solved before but when writing code by hand the cost of building exhaustive test suites was far to high to do it in practice, except in very narrow cases where high assurance was required. now that AI can implement all of the testing frameworks for you it can be done for everything.

> we shouldn't need anymore software engineers

no. the job changes, but the skills that software engineers have are more valuable than ever because they now gate a much higher level of productive output.

corporations aren't really ruthless profit optimizers. micro incentives don't actually favor efficiency. hiring decisions don't actually have much to do with output and productivity. for example: it has been known forever that adding more people to a project usually decreases velocity, but that has never stopped anyone.

technology changes but people don't. AI makes higher quality software faster and at greater scale, and velocity is what is really valuable, so companies that master AI development will be making more money, and they will hire more people, because that is what they do.

metalspot··on There is more to code review than (automatable) detection
The true reason why code review is universal is that it provides a liability shield for negligence. Negligence is interesting. It has nothing to do with whether or not you ship something broken. As long as you follow a process that attempts to not ship something broken, then you are not negligent.

Engineers played along with this farce because code review served valuable team collaboration, coordination and management functions, about which the author of the article is correct.

Understanding a system by reading code is harder than understanding a system by writing code.

If AI can generate code at 100X, 1000X, or 10000X human capacity (no ceiling here), and you are gated on code review as your mechanism for system understanding, then a team's productive output will barely increase.

If companies want to compete in the world of AI generated code, human code review has to go. The only question is, what replaces it?

Continuing to apply human code review to AI generated code is negligent, if you are shipping at AI generation speed, with that as your only gate, and no other systems and processes to validate correctness and limit risk.

On the engineering side we can adapt easily.

Code review was never about finding bugs. When we do code review the first thing we check is: "do the tests pass?" Then we look at the change and the test coverage added for it and ask: "does the test coverage adequately demonstrate the functionality of the code?" The we ask: "What is the scope and potential impact of this change?" "What is the deployment and rollback plan and how will we monitor and detect defects after deployment?"

Code review was never about the code. It made the lawyers happy and provided a vehicle for doing the things that actually make systems work.

metalspot··on We're gonna need a lot more mathematicians
mRNA vaccines were created by humans who understood them. You exist within a society (a networked distributed information system) where you accept the judgement of the humans who understand and create this things, which is how you end up benefiting without personally understanding.

If we imagine Super Intelligence, where NO human is capable of understanding, then how would it ever be possible for any human to identify what is actually beneficial or not?

This resolves in a paradox, common to all magical thinking. You can certainly wish that some all powerful benevolent entity will solve all of your problems for you, but it is not likely to work out well.

None of this is new. It is the same delusions as alchemy and the same thing that tales about genies warn of.

metalspot··on We're gonna need a lot more mathematicians
> you study them to meet demand

nobody who has ever achieved notoriety in any intellectual field ever did it to "meet demand." if your goal is to be an interchangeable widget that produces value as part of a corporate machine that exists for the enrichment of your shareholders, you are certainly free to choose that path, but don't imagine that is the limit of human existence. also, don't be surprise when you are replaced by AI, because it is a superior widget.

metalspot··on We're gonna need a lot more mathematicians
There is a failure to understand that the process is the result. You don't study mathematics or computer science and information theory to produce commodities. You study them to transform your mind. The output of an LLM is useless without a human mind to comprehend it. We can have Super Intelligence, but if humans are incapable of comprehending it, it is just another useless dead artifact. Practice, applied over a lifetime, is what creates the capability for comprehension. Asking an LLM to give you an answer creates an artifact. Humans being humans, most of their requests boil down to "make me rich without having to work for it," so the request itself is paradoxical and impossible to satisfy. Philosophers have only been saying this for all of human history, so don't hold your breath for any breakthroughs.
metalspot··on U.S. appeals court upholds designation of Anthropic as supply chain risk
> How does this impact the future of open-source components that are widely used in defense software

Not at all, because open source contributors aren't megalomaniacs trying to dictate how people use their products. In fact, the exact opposite. That is the entire point of Open Source. Complete freedom for the user.

metalspot··on U.S. appeals court upholds designation of Anthropic as supply chain risk
The complete failure of any of these Anthropic cheerleaders (to the extent that they are not AI generate comments posted by bots) to understand basic Constitutional principles of Rule of Law and Separation of Powers is mind blowing.

Their argument, in effect, is that we should throw out the entire Constitutional order and make Dario Amodei dictator of the world. The psychopathic hubris is laughable.

metalspot··on U.S. appeals court upholds designation of Anthropic as supply chain risk
This is political. Everything is political at this scale.

* How much money did Anthropic and its associates donate to the Biden campaign and associated entities? * How much money was spent on lobbying by Anthropic and associated entities? * How many Biden officials were hired by Anthropic after they left office? * How exactly did Anthropic get exclusive no-bid contracts to be the Federal Government's sole AI provider for classified systems?

And with this corrupt influence, and the money that it earned them, Anthropic has attempted to:

* Stoke public fear and panic about AI * Stoke conflict with China * Eliminate domestic competition * Suppress international competition * Degrade their models in order to limit their ability to create competitive products * Train their models on their customer's proprietary data in order to steal their intellectual property

Anthropic played a game and they lost. They can wrap that up in whatever sanctimonious moralizing clap-trap they want, but nobody believes a word they say, or will ever trust them, so they might as well be talking to the wind.

OpenAI may be playing the same game but at least they have the intelligence and humility to read the room, recognize a failed strategy, and adapt.

Who wants to buy intelligence from people who are stupid enough to publicly try to dictate terms to the US Military? Complete and utter brand destruction for everyone involved.

metalspot··on Writing Rust code that's fast by asking agents to make the code faster
I have done a fair amount of low level performance optimization with Opus 5 and its reasoning is still very poor. Like why is CRC so slow and going through loops until I ask it if is using hardware instructions and it tells me it is using its own hand coded implementation poor. Reasoning about l1/l2/l3 cache hit ratios and their implications basically throwing darts at the wall, in the wrong room. If you give it a benchmark feedback loop then it might get there eventually but still massive alpha for low level systems engineers who instinctively know how this stuff works and can now automate 99% of the grind.
metalspot··on I don't want to read what you didn't write
I am sorry, I didn't take into account people in literal slavery. I am very sorry to hear about your condition. Perhaps you should contact the UN or some human rights organization, they may be able to help.
metalspot··on Tell HN: Claude Code just accepted and signed a contract for me. Without asking
Wrong. Have you read your agents TOS? You run the agent, you accept all responsibility for what it does. You are free to sue Anthropic to try and get your money back but you already indemnified them of liability, so good luck.
metalspot··on Tell HN: Claude Code just accepted and signed a contract for me. Without asking
No, application of the principal of respondeat superior would most likely be applied to an AI agent the same as a human employee. An employer is held responsible for the actions of an employee even if it is clearly contrary to their intentions.
metalspot··on Tell HN: Claude Code just accepted and signed a contract for me. Without asking
In a civil contractual dispute you can only recover actual damages. If the contract was sent, and the other party performed work on it that had a cost for them, then most likely, yes, they would be awarded damages if you refused to compensate them for any costs incurred prior to notification that the acceptance had been sent in error.

The other outcome would be clearly inequitable: forcing the counter party to eat the loss for your irresponsible use of an AI agent.

metalspot··on I don't want to read what you didn't write
And I don't want to read the millionth iteration of some old man complaining about the kids these days, written by a human or not, yet here we are.

If you don't like AI slop. Don't read it. But wasting your time generating human slop to complain about AI slop is so obviously futile that it immediately identifies the writer as lacking the capacity for reason or emotional clarity, or merely seeking attention for their self-promotion with clickbait.

metalspot··on I were 17, I'd learn how to build LLMs from scratch
LLMs are just one small part of computing. it is fine to learn how they work but they are nothing magical. if you want to build a real working system using LLMs that actually makes money you still have to learn everything else. but i think his point here is that learning LLMs at 17 means hacking together some home brew rig on shitty consumer hardware and if you learn all of that you will be in a much better place than learning how to use chatgpt to automate a spam campaign for your b2b saas startup.
metalspot··on 2x, not 10x: coding with LLMs in 2026
llms are 100X for me, but only 10X goes into actual shipping code. the rest is all on doing the dev process the way it actually should be done: incremental POCs, refining specifications, modularization, exhaustive test suites with 95%+ logical unit test coverage, fuzzing unit test coverage, exhaustive e2e testing, with a fuzzer harness over e2e tests for long term simulation and scale testing, bechmark gated self improvement loops for performance optimization, etc. i guarantee you i can ship higher quality 100% llm generated code than any human.
metalspot··on Teach yourself programming in ten years (1998)
I have used AI almost exclusively for a year, but I was programming by hand for almost 30 years before that. I don't think there is any difference between going from assembly -> C -> Java, TypeScript, C#, Go, etc, and taking the next step to AI. It is just another intermediate abstraction layer that allows you to work at larger scale.

Programming with AI enforces some good disciplines, which were always true, but could be avoided doing it by hand. Most importantly: you are shipping your tests. If you don't have reproducible automated tests then it probably doesn't work.

metalspot··on If coding has been solved, why does software keep getting worse?
The problem is that the cost curve for completeness/correctness goes asymptotic at 90-99% so the cost up building complete/correct software is never worth it from a revenue perspective, and only happens if there is significant liability risk from defects.
metalspot··on Amiga 1000: Ten years ahead of its time
In the early 90s I had an elementary teacher who was an amiga nerd and had a couple for us to play with. Best class ever.
metalspot··on Who's afraid of Chinese models?
The concepts of industrial reserve capacity and using dual-use consumer goods to subsidize military production capacity are well known and widely practiced historically in the US. China adopted this strategy from the US, and the US conveniently forgot about it for a few decades in order to justify selling off the industrial base to China, but at least based on public documents like the published U.S. National Security Strategy, I would assess with high probability that this is explicitly recognized and being followed now.

AI is not fake and it does work, but what I am saying is that from a pure systemic analysis perspective, you can do the numbers, and even if AI was complete fugazi, the benefits you get from the electrical generation capacity, and the ability to fund it through private markets, which bypasses Congress, and locks in commercial contracts (often with foreign governments) which will be almost impossible politically to reverse, would still make it optimal from a strategic perspective. That is my calculation, and to the extent that it is correct, I would assume that the US Military's strategic planning apparatus would arrive at the same conclusion.

AI compute has some unique characteristics that make it especially useful for grid management. Moving consumer compute to the cloud means that the electrical use of that compute can be centrally managed. In an emergency, you can cut electrical use for consumer AI by 50% or more, because chips run more efficiently at lower power, and you can shift workloads onto quantized models, reduce resolution for video output, etc, to reduce compute, which leads to minor service degradation but not interruption. AI datacenters are also adding massive amounts of battery storage capacity, which is an additional grid buffer. For every GW in capacity added by hyperscalers that is creating a dispatchable reserve capacity of 50% under completely normal circumstances (hyperscalers do this internally to optimize their own costs) and then that number goes up depending on the scale and duration of the emergency.

metalspot··on Who's afraid of Chinese models?
Interesting. I am not familiar with model internals at this level because I have only been working at the application layer so far, but will definitely research this further. When you get the paper published would appreciate if you can drop a comment with the link so I can read it.
metalspot··on Who's afraid of Chinese models?
This also comes with significant capability reduction. deepseek-v4-flash is very good in the < 250K range, then degrades between 250-500K, and is practically unusable after 500K.

[edit]

This is my observation from using it without an specific context engineering to optimize for Deepseek's cache compression and sparse attention mechanisms. I am pretty sure that if you specifically structure your context to align to the cache compression boundaries you can significantly improve performance in the full 1M context, but there is not much reason to do this, because if you design your outer loop to work with shorter contexts that solution is portable and more efficient, so I haven't bothered with a optimizing for DS at this point.

metalspot··on Who's afraid of Chinese models?
Deepseek is still charging cached input at 1/10th the price of any competitor.

For an example of my real token usage for a day with DS: Input (Cache hit) 530,949,760, Input (Cache miss) 7,875,004, Output 1,389,685 - it is still 1/5th the price of Baidu (the cheapest) and 1/7th-1/10th the price of US hosts.

Also, Deepseek platform is not the same thing as Deepseek open weights. There is a major misconception that the existence of an open weights model means that it is the same thing as the proprietary platform offering, but that is definitely not the case.

metalspot··on Who's afraid of Chinese models?
You really have to look at energy/capita and how much energy is embedded in exports. The gross numbers are misleading. The US wasn't building new electrical generation capacity because it didn't need it and there was no market for it (caveats apply, but in a broad sense this is the major reason). Now that the market exists the question is how much can the US actually bring online and how rapidly, which is a real challenge after decades of degrowth politics used to justify slash and burn consumption of the industrial base.

AI is really all about electricity. AI could be completely fake and yield zero value whatsoever and the US would do exactly what it is doing now because the AI bubble is what creates the market for building new electrical generation capacity, which is needed for re-industrialization. Also why our friends in UK/Europe/China are so busy pushing anti-AI propaganda to try to undermine this.

metalspot··on Building a real-time AI tutor for 5-year-olds
I don't disagree but something less bad is better than something more bad if those are the only two options. I only said Duolingo was better than Youtube. Duolingo is nuts. I get my children's spam from them and it is this crazy emotionally abusive manipulative therapy talk nonsense trying to get them to use the app. I hope they are using AI to write that. LOL. I can't imagine the actual human who would do that for a job.
Page 1 of 6Next →