HNHacker News
TopNewBestAskShowJobs

NewEntryHN

570 karma · joined May 6, 2017

submissionscomments
NewEntryHN··on The AI Race Just Got Awkward
> So the Chinese labs have thrown a lifeline to the Western loss-making labs, and I just have no clue as to why.

Because contrarily to the author's assumption, all labs, Western or not, have sufficient skills to discover the optimizations anyway, and publishing or not is not actually that important?

NewEntryHN··on Rails World 2026 Opening Keynote [video]
Good. We need AI-bullish folks to test it boldly and then report whether it works or not.
NewEntryHN··on Navier-Stokes – Tristan Buckmaster [pdf]
This is hard to argue without a fine understanding of how much insight OpenAI had about the stab at the problem from the "public rumor" alone.

If there was any sort of coarse insight that "they're trying to solve it this way", then both those things can be true:

- The massive amount of compute from OpenAI re-discovering Buckmaster's work solely from the coarse insight (and solving the rest as well).

- OpenAI still acknowledging they basically scooped the coarse insight using compute, and they're willing to credit Buckmaster.

Buckmaster says "Concretely, what Levent and I did was to take the Cordoba and Martinez-Zoroa program, which achieved blowup results with rough forcing, and, with a great deal of help from LLMs, push it to smooth forcing and to the incompressible Euler equations".

Could that simply be the prompt they used at OpenAI? How "stolen" would the proof be in that case?

NewEntryHN··on Position: LLMs Can't Jump
You don't need any reproduction. Assuming researchers are using LLMs, you should just see the number of jumps increase as models get better.
NewEntryHN··on What's the largest software project AI can complete on its own?
> without access to the original source code

All models in the leaderboard probably have had access to the original source code in their training data.

NewEntryHN··on Prompt Injection as Role Confusion
I'm not sure I understand how important "role perception" is when following instructions from a tool call rather than the user is currently a legitimate use-case (applying steps from documentation, or shell command instructions on stdout, or really anything that can be deduced from the content of a tool call).
NewEntryHN··on How many of the 170k English words do you know?
Very easy for French speakers ahah
NewEntryHN··on You weren't meant to have a boss (2008)
Maybe shadow hierarchy are still more productive than official ones? Looks like something that wouldn't have that many meetings.
NewEntryHN··on Show HN: I made an emergency page for my family
Hmm I'm not sure why in an emergency situation accessing a webpage would be easier than making a phone call.
NewEntryHN··on It is time to give up the dualism introduced by the debate on consciousness
If the thing "outside of reality" ever reveals useful to explain anything about reality, then it becomes part of reality.
NewEntryHN··on Newton's law of gravity passes its biggest test
I think OP's question is more how could Newton's law "pass" a test any more than General Relativity would, considering that it's merely an edge case of GR?
NewEntryHN··on Newton's law of gravity passes its biggest test
Why is the article titled "Newton's law of gravity passes its biggest test" if it doesn't explain the movement more than MOND?
NewEntryHN··on Alignment whack-a-mole: Finetuning activates recall of copyrighted books in LLMs
You are comparing the fight between a p2p program and the entire music industry with the fight between the entire LLM industry and a newspaper. Notice how the order seems inconsistent.
NewEntryHN··on OpenAI ad partner now selling ChatGPT ad placements based on “prompt relevance”
Ads are for the free tier.
NewEntryHN··on SQLite in Production: Lessons from Running a Store on a Single File
It's a spectrum. Installing Postgres locally is not 100% future-proofing since you'll still need to migrate your local Postgres to a central Postres. Using Sqlite is not 0% future-proofing since it's still using the SQL standard.

If the only argument for a piece of tech in comparison to another one is "future-proofing", that's pretty much acknowledging the other one is simpler to setup and maintain.

NewEntryHN··on Post Mortem: axios NPM supply chain compromise
He says it mimicks what is described here: https://cloud.google.com/blog/topics/threat-intelligence/unc...

Which is basically phishing:

> The meeting link itself directed to a spoofed Zoom meeting that was hosted on the threat actor's infrastructure, zoom[.]uswe05[.]us.

> Once in the "meeting," the fake video call facilitated a ruse that gave the impression to the end user that they were experiencing audio issues.

> The recovered web page provided two sets of commands to be run for "troubleshooting": one for macOS systems, and one for Windows systems. Embedded within the string of commands was a single command that initiated the infection chain.

NewEntryHN··on Astral to Join OpenAI
Nothing worse than what would have happened to the Astral team if they had ran out of funding rounds without an exit...
NewEntryHN··on I don't use LLMs for programming
This assumes you always learn something new with every new program you write.
NewEntryHN··on A GitHub Issue Title Compromised 4k Developer Machines
I would not have helped. People are losing their mind over agents "security" when it's always the same story: You have a black box whose behavior you cannot predict (prompt injection _or not_). You need to assume worst-case behavior and guardrail around it.
NewEntryHN··on OpenAI raises $110B on $730B pre-money valuation
Netscape had 20 millions active users at its peak, out of 6 billions humans.

ChatGPT has 800 millions monthly active users currently, out of 8 billions humans.

NewEntryHN··on AI is not a coworker, it's an exoskeleton
This implication completely depends on the elasticity (or lack thereof) of demand for software. When marginal profit from additional output exceeds labor cost savings, firms expand rather than shrink.
NewEntryHN··on What “The Best” Looks Like
One reason you see a pareto distribution in "normal sized" teams is not solely because of competency, but because the 80% can rest on the 20% and therefore don't feel too pressed to work that much. Therefore the pareto model breaks down in 1-man teams.
NewEntryHN··on Stranger Things creator says turn off “garbage” settings
To be fair, the diction in modern movies is different than the diction in all other examples you mentioned. YouTube and live TV is very articulate, and old movies are theater-like in style.
NewEntryHN··on OpenAI declares 'code red' as Google catches up in AI race
"Software engineer complains bearing the burden of everything and concludes everything would be fixed by firing everybody except themselves."
NewEntryHN··on IQ differences of identical twins reared apart are influenced by education
Of course. IQ tests measures nothing more than the ability to pass an IQ test, which is proxied by a lot of things such as western culture, education, propensity to cram tests, etc.
NewEntryHN··on Apps SDK
Why doing it themselves instead of distributing the work to data owners?
NewEntryHN··on Apps SDK
This is not just branding, MCP is an implementation detail; the product is chatting with apps.
NewEntryHN··on Write the damn code
What's up with the "prompt refinement" business? Are folks trying to get it right with one shot?

My experience is that treating the generated code as a Merge Request on which you submit comment for correction (and then again for the next round) works fairly well.

Because the AI is bad you get more rounds than in a real code review, but because the AI is fast and in your command each round is way faster than with a code review with a human (< 10 minutes feedback loop).

NewEntryHN··on ‘Overworked, underpaid’ humans train Google’s AI
Isn't that mostly the fine-tuning phase? RLHF being cherry on top?
NewEntryHN··on Checklists are hard, but still a good thing
Would't any additional item increase safety?
Page 1 of 6Next →