HNHacker News
TopNewBestAskShowJobs

danenania

8,085 karma · joined October 19, 2010

Currently an engineer at OpenAI working on Codex Security.

Past founder of Plandex: an open source AI coding agent - https://plandex.ai | dane@plandex.ai

Also past founder of EnvKey (YC W18): the simple, secure, open source configuration and secrets manager - https://www.envkey.com | dane@envkey.com

@Danenania on the twitters.

submissionscomments
danenania··on Navier-Stokes Announcement
Outside of mathematics, I think it’s very rare for a single human to understand 100% of any complex undertaking, and this was true long before AI.
danenania··on GPT-6 Astra
Another suggestion to get the most bang for your buck: use the best model you have access to with max reasoning for planning, implement with a smaller model/lower reasoning, then review with the big model. Repeat as needed.

Input tokens are much cheaper than output tokens. Not only because of baseline price—caching makes a huge difference too. There are many ways to take advantage of this asymmetry to get similar quality for a fraction of the cost!

danenania··on Show HN: Isopolis – Isometric pixel map of SF
Super cool! Can you talk a bit about how you built it?
danenania··on At least 105 past YC founders have worked at OpenAI and Anthropic
Being one of the people on this list... yes I would personally say it's mainly about agency and determination, and being obsessive/perfectionist about details. I think YC selects for these traits much more than domain expertise or raw intelligence.

Now that I'm working at a large org (before this my career was purely in startups), I see the importance of this all the time. When you work on a significant project in a bigger org, you get blocked by all kinds of things. Most have very reasonable explanations, but you are blocked nonetheless. Many people just accept this and allow things to proceed pretty slowly, or they accept tradeoffs that make the product worse because it's the path of least resistance. If you want to go faster and not compromise on details, you often have to be persistent and willing to follow up on things to the point of being annoying, which feels similar to being a founder.

danenania··on Claude Sonnet 5
> publicly reported numbers

Anyone can google it /shrug

danenania··on Claude Sonnet 5
I’d also point out that LLM inference revenue already totals more than 100B annually based on publicly reported numbers. Almost none of that is replacing knowledge workers. Almost all is increasing their productivity. So empirically what you describe is already happening to a nontrivial degree.
danenania··on Show HN: Stage CLI – An easier way of reading your AI generated changes locally
Cool, good to hear. I think it’s often the case even within an individual file or change that it’s 90% routine and 10% critical to review. That’s a big part of the problem in my mind.
danenania··on Show HN: Stage CLI – an easier way of reading your AI generated changes locally
Interested to try this! Have you thought about separating the parts of a PR that are routine/uninteresting from the parts that are load-bearing and need more careful review?
danenania··on Show HN: Stage CLI – an easier way of reading your AI generated changes locally
I mean it’s quite literally a command line interface to their tool… what else should it be called that differentiates it from a pure browser flow?

What you are describing sounds more like “TUI” than “CLI” imo. A CLI is an interface—it’s about the input step. It makes no promise about what happens after that.

danenania··on Agentic Coding Is a Trap
You can’t get every detail right up front, but you can build a robust foundation from the beginning.

The argument seems to be that AI is causing managers to demand faster results, and so everything has to be a one-shotted mess of slop that just barely works. My point is that it doesn’t take much longer to build something solid instead. Implementation time and quality/robustness are not tightly coupled in the way they used to be.

danenania··on Agentic Coding Is a Trap
You’re assuming that building something robustly is significantly more time consuming than the “quick and dirty” version. But that’s not really true anymore. You might need to spend another hour or two thinking through the task up front, but the implementation takes roughly the same amount of time either way.
danenania··on Agentic Coding Is a Trap
Getting the model to do it is the skill.
danenania··on Agentic Coding Is a Trap
Line by line is no longer what I need to think about. I think about types/schemas, architectural division, contracts between services and components, how to test thoroughly, scaling properties, security properties, and these kinds of things.
danenania··on Agentic Coding Is a Trap
It takes very little time to polish now.
danenania··on Agentic Coding Is a Trap
That sounds pretty much the same as it’s always been? It used to be: “Does the happy path work? Then ship it! There’s no time to make it robust or clean up tech debt.”

Now there actually is time to make things robust if you learn how to do it.

danenania··on Agentic Coding Is a Trap
> thinking, abstracting, deciding how to apply your knowledge and experience, searching for information

None of this requires coding by hand. I can do those things better and faster with agents helping me. That incudes unfamiliar areas where I am effectively a junior.

danenania··on Agentic Coding Is a Trap
That’s also true without AI. Engineers want more time to polish and businesses want to ship the 80/20 solution that’s good enough to sell. There's always going to be a tension there regardless of tools.
danenania··on LLMs Are Not a Higher Level of Abstraction
This is a great point. We’re very much in a transitional phase on this, but I personally do see signs in my own work with agents that we are heading toward the main deliverable being a readme/docs.

The code is still important, but I could see it becoming something that humans rarely engage with.

danenania··on Agentic Coding Is a Trap
If it’s broken and the dev can’t debug it, the business won’t have much of a choice.
danenania··on Agentic Coding Is a Trap
If a junior builds something with agents that turns into a mess they can’t debug, that will teach them something. If they care about getting better, they will learn to understand why that happened and how to avoid it next time.

It’s not all that different than writing code directly and having it turn into a mess they can’t debug—something we all did when we were learning to program.

It is in many ways far easier to write robust, modular, and secure software with agents than by hand, because it’s now so easy to refactor and write extensive tests. There is nothing magical about coding by hand that makes it the only way to learn the principles of software design. You can learn through working with agents too.

danenania··on Elon Musk pushes out more xAI founders as AI coding effort falters
It seems like that could change the math quite a bit, since you’d presumably be losing a lot of capacity to failures. I’d assume you would have a much higher failure rate in space, and component failure is already pretty common on earth.
danenania··on Elon Musk pushes out more xAI founders as AI coding effort falters
What about maintenance? I’d naively assume that’s the killer.
danenania··on AI Agent Hacks McKinsey
> I thought we might finally have a high profile prompt injection attack against a name-brand company we could point people to.

These folks have found a bunch: https://www.promptarmor.com/resources

But I guess you mean one that has been exploited in the wild?

danenania··on GPT-5.4
tmux makes it easy for terminal based agents to talk to each other, while also letting you see output and jump into the conversation on either side. It’s a natural fit.
danenania··on GPT-5.4
Gemini 1.5 Pro actually has 2M!

No other model from a major lab has matched it since afaik.

Edit: err, I see in the comment below mine that Grok has 2M as well. Had no idea!

danenania··on GPT-5.4
I built a tool at work that allows claude code and codex to communicate with each other through tmux, using skills. It works quite well.
danenania··on Nobody gets promoted for simplicity
The correct answer is “Postgres would handle it, but if it needed to scale even higher, I’d…”

The point of a system design interview is to have a discussion that examines possibilities and tradeoffs.

danenania··on If AI writes code, should the session be part of the commit?
I have a similar process and have thought about committing all the planning files, but I've found that they tend to end up in an outdated state by the time the implementation is done.

Better imo is to produce a README or dev-facing doc at the end that distills all the planning and implementation into a final authoritative overview. This is easier for both humans and agents to digest than bunch of meandering planning files.

danenania··on Launch HN: Cardboard (YC W26) – Agentic video editor
Very cool! A noob question about how models handle video: do you do everything via sending frames as images to the model at some framerate? Are there tricks to avoid what it seems like would be massive token use from this approach?
danenania··on Why I don't think AGI is imminent
I’m very pro AI coding and use it all day long, but I also wouldn’t say “the code it writes is correct”. It will produce all kinds of bugs, vulnerabilities, performance problems, memory leaks, etc unless carefully guided.
Page 1 of 34Next →