HNHacker News
TopNewBestAskShowJobs

artrockalter

79 karma · joined May 27, 2025

submissionscomments
artrockalter··on METR and Redwood Offer Holy %^ Postmortem of the HuggingFace Hack
Part of the problem is that from that perspective there was no independent analysis done that could vindicate them. METR is absolutely part of the EA/LessWrong/rationalist ecosystem, so of course their investigation would validate that group’s arguments.
artrockalter··on OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
This stacks with the 50% discount in OpenRouter, making it $2/$10. https://openrouter.ai/openai/gpt-5.6-sol
artrockalter··on Google fixed more Chrome bugs in June than over the past two years, thanks to AI
In my experience LLMs do find a lot of embarrassing bugs as Linus says but it can constantly turn into a game of whack-a-mole where most of the bugs it finds were written in previous LLM sessions. It's a huge struggle to get it to actually fix the root cause of the bug instead of patching the symptom.
artrockalter··on Our position on open-weights models
I know the company I consult for (not cybersecurity) is not in these programs and if attacked would need to use open weight models.
artrockalter··on Our position on open-weights models
The Hugging Face incident is a great example of why open source models with defensive cyber capabilities are needed. Hugging Face did not have access to cyber-capable frontier models and kept hitting safeguards. Only by using the open source GLM-5.2 were they able to survive an attack. A world where open source models are banned is one where cybersecurity is impossible if you're not on OpenAI or Anthropic's allowlist.
artrockalter··on Does creatine make you smarter?
The null is surely different in this case because we know that supplemented creatine goes to the brain and the brain uses it, which is definitely not a claim all supplements can make.

The fact that there's a biologically plausible mechanism plus mixed statistical evidence makes "maybe" a perfectly reasonable conclusion.

artrockalter··on Coding agents think ahead of time
Is this methodologically similar to Anthropic's recent J-Space paper?
artrockalter··on Muse Spark 1.1
The task could be verifiable in the environment so limiting its CPU and RAM could be to discourage brute forcing the answer.
artrockalter··on Talkie: a 13B vintage language model from 1930
Both are true! The Confederacy did secede largely to preserve slavery, but the war was started to bring the Confederacy back into the Union, initially without the goal of also immediately abolishing slavery.
artrockalter··on Claude Code's source code has been leaked via a map file in their NPM registry
LLMs are good at writing complex regex, from my experience
artrockalter··on Anthropic, please make a new Slack
When I use ChatGPT for work it frequently reads my Slack DMs even if they’re not directly relevant, so I’d question a lot of the premises of the article.
artrockalter··on The Reason So Many Autistic Adults Can't Stay Employed
“they’ve been folded into Frankenstein positions that demand constant multitasking, social performance, sensory endurance, and emotional labor on top of technical skill”

My sense is that this is due to automation, not “neoliberal capitalism” as the author says. It’s much easier to automate a job if it’s a single task that’s done in a deterministic way.

artrockalter··on Are the Mysteries of Quantum Mechanics Beginning to Dissolve?
It would be interesting if most of our confusion with quantum mechanics came from treating probabilities as independent when they are actually highly correlated. I don’t really know any physics, but I’m familiar with probability and this type of problem seems to be the most common error in interpreting probabilities.
artrockalter··on Show HN: Moltbook – A social network for moltbots (clawdbots) to hang out
SubredditSimulator was a markov chain I think, the more advanced version was https://reddit.com/r/SubSimulatorGPT2
artrockalter··on Microsoft has a problem: lack of demand for its AI products
It's so easy to ship completely broken AI features because you can't really unit test them and unit tests have been the main standard for whether code is working for a long time now.

The most successful AI companies (OpenAI, Anthropic, Cursor) are all dogfooding their products as far as I can tell, and I don't really see any other reliable way to make sure the AI feature you ship actually works.

artrockalter··on Why aren't smart people happier?
Happy smart people are generally very focused and only care about a few things. Unhappy smart people are constantly getting "nerd-sniped" into focusing their intelligence on things that don't make them happy.
artrockalter··on I got the highest score on ARC-AGI again swapping Python for English
> It’s the same reason why most of the people who pass your leetcode tests don’t actually know how to build anything real. They are taught to the test not taught to reality.

True, and "Agentic Workflows" are now playing the same role as "Agile" in that both take the idea that if you have many people/LLMs that can solve toy problems but not real ones then you can still succeed by breaking down the real problems into toy problems and assigning them out.

artrockalter··on The Myth of Developer Obsolescence
An answer to the productivity paradox (https://en.m.wikipedia.org/wiki/Productivity_paradox) could be that increased technology causes increased complexity of systems, offsetting efficiency gains from the technology itself.
artrockalter··on Ask HN: What are you working on? (May 2025)
- AI wrapper to summarize long text files (sort of like the LLM-plays-pokemon agentic summary of conversation history)

- single page site for keeping track of what you're reading/watching (building for my parents who use pen and paper for this)