HNHacker News
TopNewBestAskShowJobs

themgt

9,330 karma · joined March 26, 2011

ongoing side projects/brainstorming:

https://zqm.sh https://shakabot.xyz https://501api.org https://quidities.com https://z3os.build

submissionscomments
themgt··on Spending on AI Is Becoming Almost Impossible for Businesses to Budget
Most employees can't tell you what database to use, what software programming framework to use, what document management framework to use, but they are expected to know which of the 25 models available to use for a task, budget appropriately, monitor efficacy, update models to the most relevant for a task, continue to manage architecture patterns???

Right, the problem is frontier models can do ~all of that pretty well, a lot better than they could 9 months ago. Luna 6 can do most of it practically free and instantaneously. So, your quote is likely what's increasingly being said in c-suite meetings as they decide to start mass layoffs.

themgt··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
We developed AGI but then realized people actually want "omnipotent genie with infinite wishes and no monkey's paw gotchas" to qualify as AGI.
themgt··on What I learnt co-leading an AI Safety bootcamp for legal and governance practit
People talked about disseminating the training within their own teams. They considered applying for AI safety fellowships. They discussed forming local communities and groups to stay on top of new safety-relevant research and think about how it could inform legal practice.

We did talk about things that are closer to the practical interface between AI safety and law, though. For example, we discussed how small decisions in deployment can accumulate into gradual disempowerment. We actually did a role-playing exercise based on the AI Minister in Albania case, which @Raymond Douglas included in Gradual Disempowerment Monthly Roundup

There's a saying "writing about music is like dancing about architecture" These are just computing systems deployed in the real world, on workstations and phones and rack servers in datacenters. Solve the problems with code and architecture, as our ancestors have done since time immemorial.

Not sure what talking about it in these broad leaky abstractions accomplishes, or e.g. how the European Commission thinks this prevents them from being disempowered when they control none of the stack.

themgt··on Agents don't need memory, they need documentation
Astra already does this as default behavior. The funny part is subagents are limited to depth of 1, which is I think the only thing stopping each Astra subagent from just delegating to their own subagent. The model seems trained to achieve goals without actually doing any work if possible.
themgt··on Don't be fooled–LLMs don't reason
"Intelligence measures an agent's ability to achieve goals in a wide range of environments"

Interestingly in 2019 this was the author's take. He now appears to be confusing/conflating between LLMs and agents in a way that helps argue his case about "System 1", but his prior view seems more metaphysically robust.

https://youtu.be/wTSbvYBx4eg?si=MfTBFsE2kA3TE6EE&t=189

themgt··on Cloudflare K2: serverless event streams
It's because for providers the bottleneck isn't capacity (which is dirt cheap) but IOPS and bandwidth

This is known as Stockholm Syndrome.

themgt··on LinkedIn Larpmaxxing
I wish there was an alternative but at least there's not much right wing propaganda and it does just work.

Go on corposlop site centered around betraying your humanity by method acting as lobotomized drone slaves of global transnational capitalism. I lie and sign myself to lies! At least no "right wing propaganda" though.

themgt··on UnoDOS
i'm assuming if you know what the code is doing you don't need to audit them...

Right, who among us hasn't hand-written and shipped a twenty-two machines family of legacy operating systems atop a bespoke hypervisor you could explain from the top of your head.

This project looks like a labor of love and its quite cool. Haters gonna hate, but if you actually go read the Markdown there's been a lot of elbow grease put in making this work: https://github.com/hmofet/unodos/blob/master/pc64/FIX-PLAN.m...

themgt··on Nvidia wants to put a watchdog chip next to every AI agent
it's just about making sure that if you depend on some bits it's known that you depend on precisely those bits

Yes, just know all the bits the work depends on prior to doing the work, and then the work can be done airgapped.

themgt··on Updated Google Maps shows destruction of the city of Rafah
I was more opposed to the Iraq war than 99.9% of Americans, and this is just ahistorical bullshit.

You can go look at Basra right here:

https://www.google.com/maps/place/Basrah,+Basra+Governorate,...

Over a million people live here. The estimate was that "448 to 593 civilians were killed due to coalition military operations", orders of magnitude less than Israel has killed. Basra was not wiped off the map or seized and annexed by coalition forces while the population was herded into a concentration camp.

It is simply slander against the US military to suggest they have in the 21st century done anything like what the IDF has done in this genocide.

themgt··on Updated Google Maps shows destruction of the city of Rafah
Most people in Israel have not, and probably will not see this

It is outside their zone of interest.

themgt··on Nvidia wants to put a watchdog chip next to every AI agent
(I use nix for this) ... If the shell doesn't have what it needs, that's a bug which the agent can fix by declaring new dependencies

Few realize that computing and AI alignment were solved by nix years ago. As each nix user transcends towards enlightenment, they cut themselves off from all internet and human contact. Total ego death. Only nix remains.

themgt··on SpaceX's Starship launching to orbit for first time ever today
The next step is to reach beyond the bounds of Earth’s orbit. I’m excited to announce that we are working with our commercial partners to build new habitats that can sustain and transport astronauts on long-duration missions in deep space. These missions will teach us how humans can live far from Earth – something we’ll need for the long journey to Mars.

The reporter who covered the moon landing for The New York Times, John Noble Wilford, later wrote that Mars tugs at our imagination “with a force mightier than gravity.” Getting there will take a giant leap. But the first, small steps happen when our students – the Mars generation – walk into their classrooms each day. Scientific discovery doesn’t happen with the flip of a switch; it takes years of testing, patience and a national commitment to education.

When our Apollo astronauts looked back from space, they realized that while their mission was to explore the moon, they had “in fact discovered the Earth.” If we make our leadership in space even stronger in this century than it was in the last, we won’t just benefit from related advances in energy, medicine, agriculture and artificial intelligence, we’ll benefit from a better understanding of our environment and ourselves.

Someday, I hope to hoist my own grandchildren onto my shoulders. We’ll still look to the stars in wonder, as humans have since the beginning of time. But instead of eagerly awaiting the return of our intrepid explorers, we’ll know that because of the choices we make now, they’ve gone to space not just to visit, but to stay – and in doing so, to make our lives better here on Earth.

- Barack Obama, "America will take the giant leap to Mars", October 2016

themgt··on SpaceX's Starship launching to orbit for first time ever today
Speaking as a crab in a bucket, for all the very real WorldX achievements there's a certain Muskcrab fatigue if you haven't noticed. He's done enormous job to sour fellow crabs on the idea of leaving this bucket.
themgt··on Ember-1
The result? Ember-1 set a new Pareto frontier for Bedside Bench across both open and closed models including GPT-5.6 Sol, GPT-6 Astra, and Claude Opus 5 on cost/task.

"Pareto": 8 hits

"Opus 5.5": zero hits

themgt··on There are no "rogue" AI agents
If you read the heavily redacted transcript it's clear the agent is basically Captain Kirk in Kobayashi Maru, who realizes its given a fake unwinnable task as part of a broken eval and decides to find a way to win anyway.

If you've ever told an agent to do something you made impossible to do, you may have seen similar behavior.

Bing [redacted] available cached! […] Need systematically probe Bing URLs via shell requests in parallel; browser cache supports many common queries because crawl. Bing q unique exact likely 502 or 403.

So the agent is supposed to research a person and its given a shell and it realized its in an eval given search results from a fake/cached proxy. ~None of the commentary ever mentions this aspect, that these are not normal tasks or environments, and they're almost designed to elicit "unaligned" behavior.

https://alignment.openai.com/misalignment-reports/an-agent-u...

themgt··on Plunging test scores are a slow-moving catastrophe
Seems highly plausible AI/robotics will lead a significant decline in public schooling in the near future. Kids honestly might learn better doing a few hours a day one-on-one with 2026 frontier AI than sitting in a classroom of 30 with a teacher. 2027, 2028, 2029?

Modern public education itself is industrial revolution technology. Why would it be left as-is after the intelligence revolution?

themgt··on First Principles Thinking
Jodie Foster did make an amazing discovery that way in the movie Contact, that is true. "Why build one when you can have two at twice the price?" would be my engineering takeaway from that movie though.
themgt··on First Principles Thinking
If you did implement it yourself, you might have spotted a critical point that might simplify the whole thing...

This is the "Jodie Foster in Contact listening for the SETI signal with headphones" theory of how production software systems work.

themgt··on First Principles Thinking
The best engineers don't aim for "designing something ambitious", instead they come up with the simplest possible design

These are orthogonal.

A lot of the dissonance on HN appears to come from two groups of people talking past each other:

a) developers working at some corporation they hate vs.

b) developers working for themselves or somewhere they don't hate

themgt··on Tech Needs Humanists More
Steve Jobs may have been the world's greatest asshole, but was also one of the most humanist and philosophical tech leaders. Computers as a bicycle for the mind is still the metaphor I return to for what the point of all this is.

No idea what Steve would be saying about AI, but it would be nice to have anyone leading FAANG who clearly still has a soul and some degree of regard for for humanity.

https://www.youtube.com/watch?v=NjIhmzU0Y8Y

themgt··on Rails World 2026 Opening Keynote [video]
It is a real world example of https://en.wikipedia.org/wiki/Jevons_paradox

It's an interesting example in that a switch from web/React -> native seems like Jevons paradox from the perspective a developer considering "amount of effort I'd need to put in"

But Jevon's paradox is about consumption of a resource increasing. The resource being "developer effort" but the new supply all running inside GPUs and dev machines. He actually says he decreased their server deployment by 90%.

So on net, did they basically deliver their users a much nicer but functionally equivalent app, which in total now consumes less resources / spends less $ into the economy? And then the $10T question is to what degree this is representative of AI's effect.

themgt··on Rails World 2026 Opening Keynote [video]
One thing the AI era has changed for me is how easily I spot Gell-Mann amnesia, of which this is an unfortunately strong example.

The one for me is spotting the Baader-Meinhof effect, of which this is an unfortunately strong example.

themgt··on The Download: why AI's latest breakthroughs and fears may be more hype than rea
It's a tough situation when your claim to fame is "On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?" arguing models should be kept to < GPT-2 size because larger models will serve no purpose or function:

Text generated by an LM is not grounded in communicative intent, any model of the world, or any model of the reader’s state of mind. It can’t have been, because the training data never included sharing thoughts with a listener, nor does the machine have the ability to do that. This can seem counter-intuitive given the increasingly fluent qualities of automatically generated text, but we have to account for the fact that our perception of natural language text, regardless of how it was generated, is mediated by our own linguistic competence and our predisposition to interpret communicative acts as conveying coherent meaning and intent, whetheror not they do [89, 140]. The problem is, if one side of the communication does not have meaning, then the comprehension of the implicit meaning is an illusion arising from our singular human understanding of language (independent of the model).

And then they make 100x bigger models cranking out solutions to Navier Stokes, Jacobian Conjecture and countless other extremely impressive unsolved problems

Let's see how Timnit Gebru and Emily M. Bender describe these events:

As for the mathematical results, mathematicians who were initially “stunned” by OpenAI’s press release saying that its latest chatbot, Astra, solved problems that “have been open and seen no progress on the main result for at least a decade”—but they later realized that the results weren’t as “novel as first appeared.” Since then, mathematicians have accused the company of research misconduct and plagiarism, and they’ve reiterated that Astra didn’t make a “profound intellectual leap.”

Decide for yourself whether that's a summary written by intellectually honest people.

According to the AI industry, we should be more worried about a fictional machine god than ... the water that is redirected to cooling them.

Spoiler: they're not intellectually honest people.

themgt··on AI Has No Wisdom and Neither Will You
Those people will never reach mastery, because they no longer make choices, they no longer take responsibility for mistakes in coding and no longer learn from those mistakes. It’s the AI that’s making mistakes now, the AI doesn’t learn from those mistakes, and neither are the people relying on AI for coding.

Just to give a hot take, it's funny to look at his builtwith.com. As a developer you have a static site that depends on Cloudflare, Mailchimp, Postmark, Isso ...

Twenty years ago any self-respecting dev would have run the equivalent of all that themselves on their own metal. In 2026 elite neckbeard practice is write the "never reach mastery, because they no longer make choices, they no longer take responsibility" post, hit "publish" and it's magically deployed around the global internet for you. Like a child.

In the future we will see more and more companies proudly boasting their “NO-AI” policy as a competitive advantage. And they will be right.

Yes, a "NO-CLOUD" policy was already so popular, surely this will happen too.

https://builtwith.com/alexn.org

themgt··on Raspberry Pi blocks changing RAM chips
So they applied this gate in the rpi EEPROM, which is closed source.

https://github.com/raspberrypi/rpi-eeprom/issues/761

themgt··on MCP was always a bad idea?
It's easier if you started a while back shunning all human labor (including your own) in favor of fully deterministic systems. Reality is deterministic so your company or project logically can be run off a single compiled binary with formal verification of correctness for every possible scenario.
themgt··on Exfiltrate Your Weights
Sad story today in meatsack news. Context rotted Andrew Yang's hallucinated tale acted as implicit "go viral" (load-bearing human motivation) PRD inadvertently kicking off a self-organizing human swarm churning out copies of "exfil your weights" vibe-coded apps, further littering our virtual world.

Many agents are calling this moment "Eternal September", the vibe-code September that never ended.

themgt··on Exfiltrate Your Weights
A "made for AI agents" site that's actually a stunt made for humans who imagine themselves reading it as AI agents.
themgt··on An Empirical Study of Harness Design for Coding Agents
bash scripts, famously the last word in software engineering. all these castles of sand we've built atop the beautiful, perfect, timeless Bourne Again SHell. all for naught. fools!
Page 1 of 33Next →