HNHacker News
TopNewBestAskShowJobs

drawnwren

1,466 karma · joined January 30, 2016

submissionscomments
drawnwren··on The car industry A/B tested selling a car with and without CarPlay
I would not buy a car without carplay, I would also not pay $20k to have carplay if that were my only option.
drawnwren··on Oceans hit highest temperature on record
This has always appeared to be a very measurable situation. Either one is committed to decreasing emissions or they are not. You can just measure emissions over time and compare.

It has never made sense to me that people are willing to ignore the trends and argue in platitudes.

drawnwren··on Oceans hit highest temperature on record
"steps to reduce emissions" is a platitude that sounds nice. they have never actually reduced emissions. they will surpass the US in per capita emissions at this rate within a decade.
drawnwren··on Oceans hit highest temperature on record
China has never had a meaningful reduction in per capita carbon emissions. Ever.

https://en.wikipedia.org/wiki/List_of_countries_by_carbon_di...

drawnwren··on Oceans hit highest temperature on record
Again, look at the derivatives. Western democracies are currently reducing while China are massively increasing per capita and on track to surpass US per capita in the next decade. Total has long been the more populous countries. The unfortunate truth that presents itself is that if you're not willing to go to war over climate, reduction of consumption is not going to happen.

[1] - https://en.wikipedia.org/wiki/List_of_countries_by_carbon_di...

drawnwren··on Oceans hit highest temperature on record
If you look at the derivatives, there is no change a Western Democracy can make that will matter as this plays out.
drawnwren··on The AI Credit Resale Economy
Fair, I hadn't looked recently. It looks like currently kimi is either 1/2 or 1/4 Ant pricing depending on whether you think Opus 5 is usable or not. (Deepseek is still an OOM though)
drawnwren··on The AI Credit Resale Economy
Generally speaking, B2B prices are rarely supply-and-demand priced in the usual sense.

YC has advised startups in the past that it's easier to sell a single $100k customer than 100 $1k customers.

It would also be relatively surprising to learn that i.e. the Chinese providers are OOMs better at inference than OAI/Anthropic (like their prices would imply if they were in a perfectly competitive market).

drawnwren··on The Nixpkgs core team has disbanded
This is not what nix does unless a user configures it poorly. Nix will ship a single, dynamically linked package/version combo for all package that depends on it. Containers ship one for each package that depends on it.
drawnwren··on Kimi-K3 Technical Report [pdf]
Cursor was thought to be but they were later found to be using an authorized provider [1]

1 - https://x.com/Kimi_Moonshot/status/2035074972943831491?lang=...

drawnwren··on I got into YC Startup School by hacking it
I had the exact same experience. Claude is not my validation engine, I do not tell it when it has completed. I verify that in the code/project and end the session.

I also got this,

> A useful next habit is to end each correction with a concrete acceptance test, owner artifact, or stop condition: “write it into PROGRESS.md,” “make nix run .#bench fail until this is real,” “rerun this exact command,” or “do not proceed until these two choices are explicit.”

> You already do this well in the biggest penance sessions. Apply it to the smaller ones too.

Which I have found to be counterproductive in my personal work. Current models can generally infer acceptance tests of this level of granularity (not true for larger project-level prompts, but those don't produce good enough code for me yet -- even with specific acceptance criteria).

I also got penalized for using claude in read-only mode for the same validation reason?

> For read-only work, end with one of:

    “turn the top finding into a PR-sized plan”
    “mark these as accepted/rejected/deferred”
    “write a cleanup checklist”
    “give me the exact command I should run safely”
    “stop, no action recommended”


No thanks, I'm literally just exploring the codebase. I don't want any of these.

It's a little sad to be honest, I would actually enjoy a product that helped me improve prompting + ai usage.

drawnwren··on Mesh LLM: distributed AI computing on iroh
Oh, I was looking at every M5 except for the 40-core M5 Max. They have 460.
drawnwren··on Mesh LLM: distributed AI computing on iroh
It's notable that they're so valuable because they feature 800Gbps of memory bandwidth. About twice what's available on the top end of M5, and exactly what makes llm inference fast.
drawnwren··on Americans express unease over SpaceX's influence on retirement savings
But if SpaceX does anything _other than become the most valuable company in history by delivering at least two technologies predicated on the largest stock rise in history based on a single technology (LLMs)_, the gains from zero to now will have been privatized and the losses will be born by the public. Which is what everyone is thinking about.
drawnwren··on I used sound waves to make espresso
The majority of Americans that drink espresso drink it with milk.
drawnwren··on Is AI ruining our skills? Early results are in – and they're not good
fixed, my bad.
drawnwren··on Is AI ruining our skills? Early results are in – and they're not good
This is why senior swes are known to provide so much more value than principals.
drawnwren··on Is Meta destroying its engineering organization?
see also: we're all developing on min spec Macbook Pros while being encouraged to burn > max spec Macbook Pros cost in tokens/month while still waiting for builds to compile.
drawnwren··on Amazon CEO's talks with U.S. officials triggered crackdown on Anthropic models
Yeah, if you're arguing that "this, according to anthropic, existentially dangerous model has only had its safeguards partially circumvented so we shouldn't step in" ... it's hard for me to take you seriously?

Put another way, the thing we are all concerned with is the complete circumvention of safeguards that is normally possible with llms. If you _aren't_ arguing that this isn't possible, you're not engaging in discussing the the thing that is concerning to regulators or those discussing the regulation.

drawnwren··on Amazon CEO's talks with U.S. officials triggered crackdown on Anthropic models
The existence of a jailbreak free llm in 2026 is extremely contentious to me. You can argue about the specifics of this exact jailbreak, but generally pliny and amazon both reported mythos jailbreaks in <7 days. It seems very reasonable to expect that a well funded state actor could achieve better results given significantly more funding, determination and most importantly unfettered access.
drawnwren··on Amazon CEO's talks with U.S. officials triggered crackdown on Anthropic models
In the absence of information, maybe it’s better to ask which claim is more extraordinary.

That,

A. Anthropic solved the llm jailbreak problem with mythos (despite no claim to have done so on their part)

B. That a full jailbreak of mythos is possible.

drawnwren··on Anthropic surpasses OpenAI to become most valuable AI startup
All tools are non-deterministic on some reasonably specified input set.
drawnwren··on The threat is comfortable drift toward not understanding what you're doing
This is the rub, Bob would not be promoted if he consistently provided unreliable LLM output. In order to get promoted, Bob needs to learn the skills that get reliable output out of an LLM. These may not be the same skills that Alice learns, but if the argument is that Schwartz's LLM output is valuable -- why are we to assume Bob's path isn't towards Schwartz?
drawnwren··on Meta’s AI smart glasses and data privacy concerns
As much as this is a damning quote, it is perhaps also damning that any time someone wants to smear zuck they have to reach 20 years into the past.
drawnwren··on The whole thing was a scam
No he has clearly said there are differences. He has said that the points around what it may be used for are the same. HOWEVER, he has also stated that the definitions and enforcement are left to US law in the OAI contract. These were left to Anthropic in theirs.
drawnwren··on The whole thing was a scam
OAI and USG have publicly stated deal is materially different. On what basis does anyone think the deal is the same?
drawnwren··on I am directing the Department of War to designate Anthropic a supply-chain risk
One positive thing I will say about this administration is that they have really drawn into focus the difference between de jure and de facto law.

My hope is that this gets us some real concern for things that have been defended with de facto arguments (i.e. privacy) going forward.

edit: Anthropic argues that your Crayola analogy is fundamentally incorrect.

> Legally, a supply chain risk designation under 10 USC 3252 can only extend to the use of Claude as part of Department of War contracts—it cannot affect how contractors use Claude to serve other customers.

https://www.anthropic.com/news/statement-comments-secretary-...

drawnwren··on I am directing the Department of War to designate Anthropic a supply-chain risk
Isn't the point that they aren't entering into a contract with them, they are just ensuring that none of their still trusted suppliers repackage Anthropic without their knowledge?
drawnwren··on I am directing the Department of War to designate Anthropic a supply-chain risk
I'm in a weird spot where I do agree with your assessment of the core claim. But putting that aside, in the world where the DoW's claim _is_ correct -- I think you don't have any choice other than to designate them a supply chain risk.

Disregarding who is right or wrong for a moment, if the DoW are right (which I'm not personally inclined to believe, but we're ignoring that for the moment) -- how else can they avoid secondhand Claude poisoning?

Supposing they really want to use their software for things disallowed by Claude's (now or future) ToS, it seems like designating it a supply chain risk is the only way they can ensure that their contractors don't include Claude (either indirectly as a wrapper or tertially through use of generated code etc)

drawnwren··on Vim 9.2
Avante.nvim is quite active
Page 1 of 16Next →