HNHacker News
TopNewBestAskShowJobs

thorum

2,727 karma · joined December 30, 2012

submissionscomments
thorum··on Microsoft account bugs locked me out of Notepad – Are thin clients ruining PCs?
AI for help figuring things out and Timeshift for when you accidentally break something. One reboot and it’s fixed.
thorum··on I miss thinking hard
> but the number of problems requiring deep creative solutions feels like it is diminishing rapidly.

If anything, we have more intractable problems needing deep creative solutions than ever before. People are dying as I write this. We’ve got mass displacement, poverty, polarization in politics. The education and healthcare systems are broken. Climate change marches on. Not to mention the social consequences of new technologies like AI (including the ones discussed in this post) that frankly no one knows what to do about.

The solution is indeed to work on bigger problems. If you can’t find any, look harder.

thorum··on Generative AI and Wikipedia editing: What we learned in 2025
I’m honestly surprised LLMs are still screwing up citations. It does not feel like a harder task than building software or generating novel math proofs. In both those cases, of course, there is a verifier, but self-verification with “Does this text support this claim?” seems like it ought to be within the capabilities of a good reasoning model.

But as I understand the situation, even the major Deep Research systems still have this issue.

thorum··on AGENTS.md outperforms skills in our agent evals
The article presents AGENTS.md as something distinct from Skills, but it is actually a simplified instance of the same concept. Their AGENTS.md approach tells the AI where to find instructions for performing a task. That’s a Skill.

I expect the benefit is from better Skill design, specifically, minimizing the number of steps and decisions between the AI’s starting state and the correct information. Fewer transitions -> fewer chances for error to compound.

thorum··on Gas Town's agent patterns, design bottlenecks, and vibecoding at scale
Agree that planning time is the bottleneck, but

> 3 days

still seems slow! I’m saying what happens in 2028 when your entire project is 5-10 minutes of total agent runtime - time actually spent writing code and implementing your plan? Trying to parallelize 10m of work with a “town” of agents seems like unnecessary complexity.

thorum··on Gas Town's agent patterns, design bottlenecks, and vibecoding at scale
Am I wrong that this entire approach to agent design patterns is based on the assumption that agents are slow? Which yeah, is very true in January 2026, but we’ve seen that inference gets faster over time. When an agent can complete most tasks in 1 minute, or 1 second, parallel agents seem like the wrong direction. It’s not clear how this would be any better than a single Claude Code session (as “orchestrator”) running subagents (which already exist) one at a time.
thorum··on Talking to LLMs has improved my thinking
I agree that LLMs can be useful companions for thought when used correctly. I don’t agree that LLMs are good at “supplying clean verbal form” of vaguely expressed, half-formed ideas and that this results in clearer thinking.

Most of the time, the LLM’s framing of my idea is more generic and superficial than what I was actually getting at. It looks good, but when you look closer it often misses the point, on some level.

There is a real danger, to the extent you allow yourself to accept the LLM’s version of your idea, that you will lose the originality and uniqueness that made the idea interesting in the first place.

I think the struggle to frame a complex idea and the frustration that you feel when the right framing eludes you, is actually where most of the value is, and the LLM cheat code to skip past this pain is not really a good thing.

thorum··on Nanolang: A tiny experimental language designed to be targeted by coding LLMs
Your other comment sounded like you were interested in learning about how AI labs are applying RL to improve programming capability. If so, the DeepSeek R1 paper is a good introduction to the topic (maybe a bit out of date at this point, but very approachable). RL training works fine for low resource languages as long as you have tooling to verify outputs and enough compute to throw at the problem.
thorum··on Nanolang: A tiny experimental language designed to be targeted by coding LLMs
Go read the DeepSeek R1 paper
thorum··on Nanolang: A tiny experimental language designed to be targeted by coding LLMs
Developed by Jordan Hubbard of NVIDIA (and FreeBSD).

My understanding/experience is that LLM performance in a language scales with how well the language is represented in the training data.

From that assumption, we might expect LLMs to actually do better with an existing language for which more training code is available, even if that language is more complex and seems like it should be “harder” to understand.

thorum··on The unbearable frustration of figuring out APIs
I remember reading and hearing similar rants from programmers 15 years ago, long before LLMs. The author kept going and figured it out, and probably got some pride and enjoyment from finishing the project in spite of the frustrating moments. That’s what learning to code has always been like.
thorum··on Influencers and OnlyFans models are dominating U.S. O-1 visa requests
I would disagree. If you have no class in private, you have no class.
thorum··on Influencers and OnlyFans models are dominating U.S. O-1 visa requests
> Hollywood had a ton of issues but it at least had some... class?

It looked that way because they had media training and their public personas were carefully managed, with staged interviews and media appearances. Behind the scenes, it’s a different story.

Influencers are rewarded for seeming authentic. Mr Beast coming across badly in a traditional TV interview just makes his audience think he’s more real.

thorum··on AI generated music barred from Bandcamp
Can’t imagine this policy lasts more than a year or two given the rate that AI tools for music are improving. Once the tech can reliably create high quality dry stems of instruments, backing tracks etc. and automate professional-sounding production work (which most musicians do not currently have access to) everyone is going to be using it even if they won’t admit it publicly.
thorum··on Anthropic blocks third-party use of Claude Code subscriptions
AI labs are not charities and there is no way to make money offering unlimited access to SOTA LLMs. Even as costs drop, that will continue to be true for the best models in 2027, 2028 etc. - as demonstrated by the fact that CPU time still costs money. The current offerings are propped up by a VC bubble and not sustainable.
thorum··on LMArena is a cancer on AI
My favorite is LLM-as-judge with a detailed rubric as discussed here: https://www.dbreunig.com/2025/07/31/how-kimi-rl-ed-qualitati...
thorum··on LMArena is a cancer on AI
Aside from Meta is there any reason to think the big AI labs are still using LMArena data for training? The weaknesses are well understood and with the shift to RL there are so many better ways to design a reward function.
thorum··on X blames users for Grok-generated CSAM; no fixes announced
“Send screenshots of this conversation to his mother” might be more effective.
thorum··on X blames users for Grok-generated CSAM; no fixes announced
They don’t seem to have taken even the most basic step of telling Grok not to do it via system prompt.
thorum··on A website to destroy all websites
The real trend is toward personalization on the user’s side of things. Instead of interacting directly with a website, your web-browsing agent will extract the parts of the website you actually care about and present them to you in whatever format, medium and design style you prefer.
thorum··on Show HN: Superset – Terminal to run 10 parallel coding agents
How are people productive using 10 parallel agents? Doesn’t human review time become a bottleneck?
thorum··on A Proclamation Regarding the Restoration of the Dash
The problem isn’t the em dashes, it’s the overuse of em dashes. Same for all the other ChatGPT-isms - they’re fine when used occasionally for effect, but there’s no variety. It’s always the same punctuation, same grammatical structures, same rhetorical moves, same paragraph lengths... That’s not what writing is supposed to be like and it becomes very grating after a while.
thorum··on Everyone in Seattle hates AI
People hate what the corporations want AI to be and people hate when AI is used the way corporations seem to think it should be used, because the executives at these companies have no taste and no vision for the future of being human. And that is what people think of when they hear “AI”.

I still think there’s a third path, one that makes people’s lives better with thoughtful, respectful, and human-first use of AI. But for some reason there aren’t many people working on that.

thorum··on Fighting the New York Times' invasion of user privacy
> Our long-term roadmap includes advanced security features designed to keep your data private, including client-side encryption for your messages with ChatGPT. We believe these features will help keep your private conversations private and inaccessible to anyone else, even OpenAI.
thorum··on A Fond Farewell
This press release has a bit more explanation:

https://www.farmersalmanac.com/end-of-an-era-farmers-almanac...

> This decision, though difficult, reflects the growing financial challenges of producing and distributing the Almanac in today’s chaotic media environment.

thorum··on ChatGPT terms disallow its use in providing legal and medical advice to others
IANAL but I read that as forbidding you to provision legal/medical advice (to others) rather than forbidding you to ask the AI to provision legal/medical advice (to you).
thorum··on Do you know that there is an HTML tables API?
createElement(‘tr’) and table.appendChild(row)
thorum··on GenAI Image Editing Showdown
Actual link seems to be: https://genai-showdown.specr.net/image-editing
thorum··on Beliefs that are true for regular software but false when applied to AI
The source code is not the LLM. The LLM is billioms of random floating point numbers that somehow encode everything the model knows and can do.

The ML field has a good understanding of the algorithms that produce these floating point numbers and lots of techniques that seem to produce “better” numbers in experiments. However, there is little to no understanding of what the numbers represent or how they do the things they do.

thorum··on Microsoft is plugging more holes that let you use Windows 11 without MS account
Am I the only one who has simply said “no thanks” to Windows 11?

There were hints of where Microsoft was heading in Windows 10, but at least a lot of the worst “features” could be disabled.

I find 11 just completely unacceptable software to run on any system I own.

← PreviousPage 2 of 15Next →