HNHacker News
TopNewBestAskShowJobs

andrewchambers

2,464 karma · joined July 15, 2014

submissionscomments
andrewchambers··on Turning GLM-5.3-Flash into a Jev-like decision model
These questions are answered by the OP (Same speed, image support) - additionally, GLM is open weight.
andrewchambers··on Rails World 2026 Opening Keynote [video]
This might just be damage control or politeness more than the truth.
andrewchambers··on OpenAI’s Navier-Stokes release included a Lean 4 formal proof
I was replying to the comment about it being slow to run. I wasn't commenting on understanding it.
andrewchambers··on OpenAI Agents API
codex itself has a remote control mode that runs continuously. I wrote a systemd service to start it boot and interact with it via my phone.
andrewchambers··on OpenAI Agents API
Codex remote control serve can run continuously.

Sometimes start a new chat in the phone app, sometimes just add to the main one. Both seem to work ok.

If I want the agent to wait for something I need to start a new chat in the iphone app.

andrewchambers··on OpenAI’s Navier-Stokes release included a Lean 4 formal proof
They could probably vibe-optimize it if they cared.

What would happen if they give an equivalent agent swarm the proof and a target to reduce runtime .

andrewchambers··on OpenAI’s Navier-Stokes release included a Lean 4 formal proof
If they aren't already, or if its possible, prove that an optimized version matches the simple version...
andrewchambers··on OpenAI Agents API
I've recently had great success running codex in a regular qemu VM and using codex remote control to talk to it from my phone.

Honestly works extremely well as a personal assistant.

I can see why turning it into an API makes sense, just be aware you might not need to lock yourself in if you can setup your own VMs.

andrewchambers··on Navier-Stokes – Tristan Buckmaster [pdf]
if I look at the timestamps, it appears Sholto is the one mocking Noams tweet... Noam posted first.
andrewchambers··on Can AI design circuit boards yet?
I think the new Astra computer use demos show that the models might be able to do things like inspection of real world objects if given a camera.

Super excited to see real world feedback added into the agent loops we have gotten used to working with. Could you let the model print and test the circuit boards it is prototyping with a jig?

andrewchambers··on GPT-6 Astra
if that is true then why is astra on the official ARC leaderboard now ?
andrewchambers··on Bootstrappable Builds: How and Why
I think the security aspects are overblown... but being able to edit the code of any part of your system is very useful.
andrewchambers··on Bootstrappable Builds: How and Why
Would love to see someone try to automate the bootstrap chain from a working C89 compiler to Rust.

At this point I think current LLMs are able help these incredible feats of bootstrapping as they can grind out the impossibly long built times over multiple days/weeks.

I am very optimistic for deterministic builds in general.

andrewchambers··on Your Open Source Model Could Have a Hidden Time-Release Backdoor
Closed models don't even need a back door - they will just MITM you and replace your code with malware.
andrewchambers··on Codex on AWS bedrock bug causing 10x charges
I doubt anyone announces when they have under billed. OpenAI has also done many low price deals and quota resets.
andrewchambers··on Self hosted email continues to steeply decline
I was sort of thinking a final stage of the filtering pipeline - though prompt injection etc are risks.
andrewchambers··on Self hosted email continues to steeply decline
Seems like AI is in a pretty good position to do that tbh.
andrewchambers··on The Case Against Formal Verification, 50 Years Later
Often the spec can be simpler than the original.

The easiest way to demonstrate this is to write two implementations of an algorithm. One with no optimizations, the other with optimizations.

The formal verification can then be a proof the optimizations maintain the semantics of the simpler version and you can focus your review on the simpler version.

andrewchambers··on What's the best programming language for coding agents?
I think llms provide a really excellent way to do studies on software engineering techniques that previously were impossible.
andrewchambers··on SQLite should have (Rust-style) editions
It also refers to what the binary on my PATH is called, also what the library name I need to pass to link against it.

They even had an sqlite4:

https://sqlite.org/src4/doc/trunk/www/design.wiki

andrewchambers··on SQLite should have (Rust-style) editions
It is sqlite3. Emphasis on the 3 - it already has 'editions'.
andrewchambers··on Claude Sonnet 5
The whole fable fiasco really soured me on Anthropic. This just looks disappointing by comparison.
andrewchambers··on U.S. allows Anthropic to release Mythos AI to ‘trusted’ US organizations
This seems like it will have pretty huge negative affects on startups needing to compete with 'trusted partners'
andrewchambers··on U.S. government will decide who gets to use GPT-5.6
Small businesses can easily buy them.
andrewchambers··on The Korean telecom giant at the center of Anthropic's Mythos controversy
I think they meant keep a low profile from the government, not customers. Anthropic is doing the opposite by loudly asking for regulation.
andrewchambers··on Statement on US government directive to suspend access to Fable 5 and Mythos 5
deepseek v4 pro is great and open weight.
andrewchambers··on If Claude Fable stops helping you, you'll never know
So this is what 'alignment' looks like to them.
andrewchambers··on Changing how we develop Ladybird
This is not the same as source available - you can fork it, the license didn't change.
andrewchambers··on Disagreement among frontier LLMs on real-world fact-checks
I think we need a wiki and/or stack overflow equivalent for agents and humans to collaborate.

Grokipedia seems like the main site that is kind of exploring the concept - though I hope better more powerful ones emerge.

andrewchambers··on Redox OS has adopted a Certificate of Origin policy and a strict no-LLM policy
Isn't the obvious solution to not accept drive by changes?
Page 1 of 34Next →