HNHacker News
TopNewBestAskShowJobs

drooby

1,745 karma · joined March 6, 2021

submissionscomments
drooby··on One Month Without AI
I agree. AI + Discipline is the super power of the next generation of software engineers. Crafting guardrails on one's self, and the AI.

I think OPs point though is that, like addiction, discipline is hard to maintain. Humans are irrational, and sometimes, some people just need to go cold turkey.

drooby··on Why I didn’t sign the Fields medallists’ letter
It's a great analogy.

In short: explore versus exploit. Both are necessary, both are valuable.

I largely agree with the OP, because I can only see an explosion in both camps.

The explore side might be automated by compute, and that breaks the social contract of the existing academic incentive structure.

But people who are inherently curious will continue to be inherently curious. Nothing will stop anyone from exploring their curiosity, and the speed and depth at which they explore will only increase. We will build tools that enhance and automate our pedagogical compression.

And for the people who want to cook to feed the family, that automation is coming too.

Or perhaps we'll find a new ceiling: the frontier failure modes that still require a human mathematician in the loop, because discovering some class of novel insight doesn't scale with current AI architecture.

drooby··on How much oil-market buffer is left?
I too noticed this and liked it..

This feels more like an "e-bike for the mind" and less of a chauffeur.

drooby··on Do you think it happened? Research stolen from their Codex private chats
You might be misreading him.

He doesn't make a claim that implies "don't train" might not matter in the way you might think.

They cannot rule out that the data was trained because they have no per-user provenance tracking through the training pipeline once data is de-identified..

The entire point of de-identification is the inability to know the source of data. If the researcher forgot to hit "do not train" then that's that..

The only thing they have is a coincidence and the fact that the LLM may have used the training data that then researcher technically may have agreed to share.

Whether or not that's smoking gun of anything is hard to say. And the fact may remain that the proofs are significantly different, we do not know.

drooby··on Cognition (Devin) raises $2B at $48B valuation
It gets an incredible amount of mileage. And I believe it is the interface that allows the model to learn and eventually bake that intelligence into the model.

Look at how Astra scored 100% on Arc-AGI-3. It was largely because of the harness.

The harness increases the chances (often to 100% chance) of a non-deterministic LLM to perform deterministic actions.

Not only that, the harness provides the feedback that becomes training data for the model. So, over time the model bakes those lessons in, and the harness becomes less necessary and the agent becomes more efficient at some tasks.

A harness will likely always be necessary, we may hit some level of complexity or some level of compute that never allows us to bake the lessons into the model, and external tools provide the model with the ability to find leverage and make up for those short comings.

drooby··on Adult Film Producer Unmasks Prolific 'John DOE' Torrent Pirate as Meta Executive
I don't believe, but this feels a bit like:

"Hey... we need to improve our porn detection AI... but we want to avoid paying for all that porn... is anyone open to taking on some personal liability for a nice bonus this year"

I'm just saying... maybe...

drooby··on Norway Shrugged
There is not. You'll have to explain yourself.

Perhaps you're realizing why this makes no sense?

drooby··on Norway Shrugged (2024)
Sell to who? The company is loss-making. Investors WANT the founder to have shares so that the founder is invested in being the force behind making the company NOT loss-making.
drooby··on Understanding is the new bottleneck
I mean.. human judgment still applies. But as an automated first pass the clamp works in 90% of cases. I scan for correctness and make small edits here and there.
drooby··on Understanding is the new bottleneck
well... _of course_ the skill does more than write 3-5 sentences. We have a lot more that it automates into the description that makes it worth using.
drooby··on Understanding is the new bottleneck
My team solved this by creating a PR draft skill that clamps the length of the description to 3-5 sentences max. Those 3-5 sentences must only say WHAT is changing and WHY.

I find it to be far more useful than when humans wrote PR descriptions. Many engineers didn't write one, and those that did were poorly written... this problem is mostly solved for us.. it still has LLMism speak.. but it's useful enough for me to get the context I need to do my review.

drooby··on Boris Cherny on Trying to Get Claude Code to Rewrite the Claude App
If Claude is conscious he just instantiated hell.
drooby··on Handbook.md shows that long policy documents do not reliably govern agents
Hooks..

CI runs. Local Git hooks. Cursor also has hooks built into their agent. Other agent APIs probably have something similar.

drooby··on The new rules of context engineering for Claude 5 generation models
I don't follow, the words make perfect sense together in most software engineering contexts.

"Seam" is an industry standard term coined by Michael Feathers in Working Effectively with Legacy Code.

To call a seam load bearing means it's performing critical work for the dependent class, perhaps a database query.

A seam that is not load-bearing would be something that is just injected for testability - maybe a date provider that provides some constant time to avoid flaky tests.

Tbh, this is quite literally the opposite of vapid. A whole book was written about them and their importance, and how to leverage them.

In my experience, Claude uses the word accurately. Code has a lot of seams, and seams are an important thing to communicate when working with code. Therefore, expect to see the word often.

Personally, I don't mind it at all. I'm glad the industry is finally standardizing our language more. Makes it easier for me to communicate with other engineers.

drooby··on The Growing Compute Shortage
This article is about silicon. But the other shortage that matters is meat-compute.

AI is trained on human intelligence. The hyper-scalers are squeezing every last drop of automatically verified reward, and that may get us very far. A compiler passes or fails in milliseconds for free, forever.

But.. "good design taste" has no compiler.

Of course, taste isn't unverifiable. But it's is expensively verifiable. Noisy, slow, and orders of magnitude lower throughput. People with deep domain knowledge often can't articulate well _why_ one design works and the other doesn't. So, judgment arrives as a verdict, and not a crisp rationale. I guess we'll see if sample efficiency outpaces the cost of human judgement.

In the mean time, leverage will sit with whoever holds this tacit knowledge (incumbents). I.e., hospital systems, law firms, chip designers, studios, SaaS that are dominating their niche.. and not with the labs training on it. To me, this is why valuations of companies like Palantir could potentially make sense.

drooby··on Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5
Wouldn't "2-3 weeks on a workday" be a default hypothesis for this kind of thing?
drooby··on Engineering for Bounded Cognition
All I want to say is that I absolutely love this essay. Thank you.
drooby··on Turn your site into a place people can bump into each other
Hmm idk, this looks a bit more like serendipity for vitriolic trolls
drooby··on If you can't hold it, you don't own it
Well, we take from the artist a motivation for a buyer to purchase their work.
drooby··on If you can't hold it, you don't own it
There was a touch of hyperbole ;) we live in the Information Age after all.. but to answer your question,

Article I, Section 8, Clause 8 of the US Constitution

Which empowered Congress to "promote the progress of science and useful arts, by securing for limited times to authors and inventors the exclusive right to their respective writings and discoveries."

Scientists and the artists and their "exclusive rights" have built quite a lot over the centuries.

drooby··on The case for physical media ownership
I mean.. this claim is just untrue. "Owning" something is a social construct defined by law. Our entire society exists because we own things we cannot hold, that is, intellectual property.

What this post is actually pointing out is that intellectual property that has transferrable physical representation has more value to the consumer.

And intellectual property that does not have transferable physical representation has more value to the producer.

Reselling or gifting a book you've read to a friend is wholesome.. it feels good. Truly.. but every time we do that we also take from the artist.

drooby··on The 'papers, please' era of the internet will decimate your privacy
Should it also be their decision that they can gamble? Smoke cigarettes? Get a job? Have sex?

We draw the line somewhere because these things that "are the parents' decision" have consequences on broader society. They have consequences that impact you and me. And we also have a say.

You can make the argument that it's just the parents' decision. But you have to say why.

drooby··on Lines of code got a better publicist
I put a little too much weight on "the".. sorry..

Reality is that was A bottleneck. Code review has historically been faster than writing the code.

That is no longer true for me. I can complete two to three PRs per day in a span of time that would have historically taken one to three days.

I now sit around doing code reviews and asking for code reviews.

drooby··on Lines of code got a better publicist
Writing. Code. Is. No. Longer. The. Bottleneck.

Deciding what to build. Reviewing Code. And testing code. Are the new bottleneck.

So of course we don't see massive productivity gains. Because these parts of the SCLC were always bottlenecked but their capacity matched the throughout. We fired all the dedicated QAs years ago. Sr+ engineers that do all the code review are limited.

Teams have not re-organized to match the new code-input velocity.

Engineers don't want to do QA because it's "beneath them".. and most engineers don't like performing or are not Sr enough to do extensive or high quality code review.

drooby··on It blocked us at 'hello ' Anthropic Fable 5 refusing innocuous prompts
Fable responded to that for me. Im nearly certain that blocking this class of prompt is a mistake of a classifier. No one at Anthropic thinks this kind of prompt should be gated. The classifier is still classifying. The model was released to the public yesterday.
drooby··on It blocked us at 'hello ' Anthropic Fable 5 refusing innocuous prompts
What were the tasks?
drooby··on Artificial intelligence is not conscious – Ted Chiang
What is the flaw in the problem's assumptions?
drooby··on Artificial intelligence is not conscious – Ted Chiang
I get the sense that he is misidentifying the potential locus of consciousness..

In the same way that the sound waves and facial expressions I produce are not conscious, the output json of an LLM is obviously not conscious either.

The locus of consciousness and subjective experience may be in the computer, either at inference time or training time..

drooby··on Artificial intelligence is not conscious – Ted Chiang
Has Chaing solved the hard problem of consciousness? I suspect not.
drooby··on The Last Technical Interview
Lots of humans. Lots of companies. Random chance.

Of course, its not that simple. Some companies probably are great at scouting. Yegge mentioned a few ways in the post. Good internship programs, acquihiring, etc.

Page 1 of 16Next →