HNHacker News
TopNewBestAskShowJobs

letmevoteplease

1,026 karma · joined August 14, 2018

submissionscomments
letmevoteplease··on Has Violence Against Teachers Become Accepted by Society?
27 unarmed black men were killed by police in 2019 according to the Mapping Police Violence database. (Despite this, 53.5% of people identifying as "very liberal" estimated that 1,000 or more unarmed black men were killed. Only 15.71% correctly estimated "about 10."[1]) There are 800,000 police officers in the United States, so I am not sure you can say they generally enjoy killing unarmed black men based on a few dozen deaths a year.

[1]https://research.skeptic.com/content/files/2025/02/Research-...

letmevoteplease··on 'That's so AI ' What gen Alpha's biggest insult tells us
Most of them are using AI for all their schoolwork while performatively hating it because "it wastes water" or some other half-understood meme.
letmevoteplease··on Claude discovers a novel enzyme system with CRISPR-like repeats
archive.is has a snapshot from 18:47, showing the link was indeed there a few hours ago. https://archive.is/XM0Nw
letmevoteplease··on OpenAI GPT–6 Astra breaks Enigma message that has resisted solution since 2005
This was done by an OpenAI subscriber, not an employee, so Astra would not have had access to OpenAI's massive compute for brute forcing. The scripts it wrote presumably ran on the computer of the customer. (ChatGPT can run scripts on OpenAI's servers, but it has a 45 second execution limit.)
letmevoteplease··on Show HN: Scry, programmable internet search w/ congestion pricing
Cool but please rewrite the text. Ctrl + F "land" gives 29 matches.
letmevoteplease··on How good are frontier models at physics?
This study appears to be evidence against that: the model failed the benchmark but arrived at the correct answer.
letmevoteplease··on Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra
> Altman was caught in previous attempts trying to game benchmarks

Sounds like something you just made up, or maybe you read it on some other Reddit/HN post and started repeating it because it aligned with your biases.

> does anyone believe that he's found his moral compass and decided to stop exploiting as much as he can get away with?

I don't think "OpenAI" is equivalent to "Sam Altman." I think if OpenAI was intentionally "benchmaxxing" purely for marketing purposes that information would leak, because OpenAI is full of good-faith researchers (although it can be difficult to avoid overfitting even if you're actually trying to improve the model's general abilities)

And lastly I think anyone can actually try Sol themselves and see that's it a good model, or if that's too subjective, it is clearly better than the previous version. The benchmarks are reflecting actual progress and anyone can verify this themselves.

letmevoteplease··on DeepSeek v4.1 Flash
I also like DeepSeek, but I'll note their stated goal is to develop AGI, and the founder (already China’s fifth-richest person) has stated, "I believe the business opportunities here are large enough-if the AI era will produce many trillion-dollar companies, I think we will be one of them."[1] These are not humble ambitions.

[1] https://liangwenfeng.art/ch11.en

letmevoteplease··on More questions about whether researchers can trust OpenAI with unpublished math
You quoted the OP saying "in this case too they didn't actually have the solution" and responded with the totally unrelated, "Given the size and recall of the biggest models, it's not unreasonable to assume that a single pertinent conversation would make it into the training data."

Neither of the researchers insinuating that their ideas were trained on had the actual solutions. This means the model could not have "stolen" the final solution from their data. At most, it could have built upon their work in the same it builds upon any other training data, though that is also questionable speculation.

>They could 100% definitely say no, if they know they did not train on user data.

No one anywhere has claimed that "OpenAI does not train on user data." OpenAI has always said that it trains on user data.

>They immediately started racing to a solution after one researcher enquired about whether they are training on their conversations.

They started racing towards a solution after they heard (incorrectly) that Anthropic had a solution; I agree this is poor sport but the "after one researcher enquired about whether they are training on their conversations" claim is false. The enquiry happened after OpenAI had obtained the solution.

letmevoteplease··on On the Navier–Stokes Millennium Prize Problem
You are confusing ideas here. No one except OpenAI had a solution to Navier–Stokes. Buckmaster and Alpöge had a solution for the forced Euler problem, which they arrived at largely using LLMs (Claude and Codex). Buckmaster implies (but does not explicitly accuse, since he has no evidence) that training on his prompts had some influence on OpenAI's result. This seems unlikely to me but is not impossible. However, in either case, the solution was found due to an LLM. Of course the LLM built on past human work, but "plagiarism" is not sufficient to account for the distance between the papers of Martínez-Zoroa, or the prompts of Buckmaster, and the final resolution.
letmevoteplease··on On the Navier–Stokes Millennium Prize Problem
This is how every conspiracy theorist thinks: my enemy is Bad, and if they did a Bad thing, it would be Good for them, therefore they obviously did it. No evidence needed other than "motive" + my enemy is evil. But even if your enemy is evil, in this case, they would be fools to take the legal risk of violating their contract for the minimal upside of a tiny bit more training data (and fools to assume this would not be exposed in a large organization). So you need to assume your enemy is both evil and remarkably stupid.
letmevoteplease··on Recreating Minecraft Is Not a Benchmark
It reads like it was generated with AI and then edited to have poorer grammar.

"A launch is a first impression and first impressions are marketing, that’s why you’re always bound to be shocked - the shock was scheduled."

"And if a model scores well but keeps failing at your work, that is actually a gap that deserves investigation."

That's Claude speaking.

letmevoteplease··on Statichost.eu – European static site hosting
If they stated the storage limits explicitly, people would just go right up to the edge. Better to leave it undefined and assess compliance afterward.
letmevoteplease··on US data centers tripled annual water consumption to 17B gallons
I think most of the people who act like this are secretly using ChatGPT like a Republican on Grindr.
letmevoteplease··on To become a better writer, read as much as you can
“Read, read, read. Read everything -- trash, classics, good and bad, and see how they do it. Just like a carpenter who works as an apprentice and studies the master. Read! You'll absorb it. Then write. If it's good, you'll find out. If it's not, throw it out of the window.” - Faulkner
letmevoteplease··on WorldClaw Agentic 3D open-world generation at scale
Even if 99% of generations are trash, why does that matter? Just throw out the failures and keep the good ones, like any writer. Here is the text of the winning story: https://granta.com/the-serpent-in-the-grove/ If you've ever used ChatGPT to generate stories, you will find the style unmistakable. It's cloying to me, but it did manage to fool some stuffy literary judges.
letmevoteplease··on How Claude marks AI-generated content
It is of course a stupid regulation, but the upside is that it will probably accelerate growth in usage of open models that are not adversarial towards the user.
letmevoteplease··on xAI, SpaceX, and the Race for AI Buildout
Look at this statement from the article:

>These racks are needed to keep up with the ballooning requirements of each new "frontier" (a buzzword label with no actual inclination of performance of improvement) model developed, which everyone bought into the AI hype will immediately hop to because it's all novelty over utility.

This is a mind-blowingly clueless assertion. This is not a person capable of rationally assessing facts. At best this kind of writing is interesting as an artifict of human delusion and cognitive bias.

letmevoteplease··on Third-party cyber evaluations involving OpenAI models
Yup. OpenAI: "the name of the fictional target for the CTF challenge unintentionally coincided with a real domain." Anthropic: "the fictional target company chosen by our evaluation partner shared a name with an active website domain name." So it sounds like the same website got hacked by both GPT and Claude.
letmevoteplease··on Investigating three real-world incidents in our cybersecurity evaluations
This does not make sense. Did you read the article? They were not trying to "catch" it accessing the internet. It did not escape. A partner accidentally left the connection to the internet open.
letmevoteplease··on Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident
Have you considered that there are reasons to do things beyond financial incentives? This incident is obviously very interesting, particular to the type of hacker employed by Hugging Face.
letmevoteplease··on Be skeptical of OpenAI's rogue hacker agent story
Go ahead and be specific about this lie.
letmevoteplease··on Be skeptical of OpenAI's rogue hacker agent story
Totally evidence-free speculation presented as fact. The average Hacker News thread about AI feels like reading /r/conspiracy.
letmevoteplease··on Codeberg Bans Cryptocurrency Projects
Looks like another project that's gonna purity spiral into People's Front of Judea vs. the Judean People's Front weirdos sniping at each other.
letmevoteplease··on Apple defeats liability for not scanning iCloud for CSAM
This is not client side. Images just sitting on your Android phone are probably safe (although Google could push an update at any time). Your images get scanned when you back them up, send them over RCS, etc. I have even seen criminal cases originating from reverse image search - anything that touche the servers of the big tech companies, except apparently Apple, will be scanned using questionable AI and against a secret list to Protect the Children.
letmevoteplease··on Donation Controversy
A mistake from a public relations perspective, but I think taking an explicit stance against deplatforming and self-censorship is consistent with Mullvad's stated mission.
letmevoteplease··on Xiaomi-Robotics-1
"the future where robots do everything" and "unemployment" are the same thing
letmevoteplease··on Holes
The story of the hand-dug well: https://www.mybrightonandhove.org.uk/places/utilities/woodin...
letmevoteplease··on We’re making Bunny DNS free
I think you're confusing this with the more classic "it's not X, but Y" trope. That sentence is a comma splice that I'd expect LLMs to avoid by default.
letmevoteplease··on Hyundai buys Boston Dynamics
Not sure how to square this post with recent headlines like "SoftBank posts $46 billion gain at Vision Fund driven mainly by massive OpenAI bet".
Page 1 of 5Next →