HNHacker News
TopNewBestAskShowJobs

PaulStatezny

629 karma · joined August 2, 2013

Software engineer
submissionscomments
PaulStatezny··on Show HN: TinyAIArena watch AI agents battle it out
The rules:

* Goal: be the last fighter alive.

* Turns: each round every fighter takes one turn. Turn order is randomized every round.

* Actions: Move one cell up/down/left/right, attack an adjacent enemy for 15–24 damage, or wait. An action consumes 1 AP.

* Rocks/Obstacles: 4 random impassable cells.

* Power-ups: Gold +1 AP per turn.

* Kills: the killer gets +1 AP per turn, and heals 50 HP (no over-heal).

(Unclear while watching replays, found in README.)

PaulStatezny··on Claude Opus 5.5
Thanks for spelling out what the original comment was implying.

I find it bizarre how intensely a bunch of these child/grandchild comments are criticizing the notion that people would even think to analyze the meaning behind the words.

Hacker News has always had a unique culture in which thoughtful discussion is basically the main goal, and it's intentionally incentivized in numerous ways. It's been my experience that any thoughts added to a post's conversation are seen as valuable as long as they are thoughtful and seeking to understand.

So these comments are clearly coming from a place that's antithetical to HN's culture. What that in mind, it seems likely to me (Occam's Razor) that these comments are either:

1. Astroturfing: Claude employees acting like everyday folks, secretly trying to shift public opinion.

2. AI cult mindset: "AI is humanity's salvation; how dare you have perspectives outside of those accepted by the cult."

Am I missing another likely option?

To bolster my point, right now we're posting on the top top-level comment, meaning a majority of active HN users find it to be a great addition to the conversation. Commenting to shut down the discussion is a red flag.

PaulStatezny··on Learning Programming in an Age of LLMs
Agree and disagree.

There have always been software companies that care about quality, and those that don't.

Many who don't care about quality exist because their products are forced onto their users. (Due to footholds from enterprise relationships, regulation, etc.) I bet slop will abound in these kinds of companies, but their codebases and products were already terrible anyways.

---

But research reliably shows users do care about things Just Working™ and feeling polished. With few exceptions, if software feels at all buggy or doesn't look visually amazing, you won't acquire/retain that many users.

Natural selection will teach hard lessons to the industry. Customers will notice things feeling "off" on products where AI slop is allowed to abound, and they'll flee to companies with sane approaches.

A sane approach: Humans actually guide the direction of the code which means they have to understand + review the code and course correct bad decisions. This doesn't mean agentic coding goes away, but it means this mad rush for insane velocity goes away.

---

Compare vibe-coded apps you've interacted with against world-class polished apps like Spotify, Gmail, Slack, etc. Those apps aren't obviously showing signs of AI slop, because the organizational structure is in place in those companies to prevent engineers from just throwing slop over the fence. Those engineers are doing agentic coding but are being forced to go at a sustainable pace.

The industry will eventually be forced (by the reality of business results) to recognize that this is the only approach that will lead to success.

PaulStatezny··on How to write an effective software design document
I think your framing is fair here. But I'd like to offer an even more complicated/nuanced take:

Designing in a group can be very difficult, and doing it well is a skill set that most people don't naturally have.

I think this explains your point 1. Why do people view designed docs as pointless? Because they really don't have a vision or model for what and effective and healthy collaborative design process would look like.

PaulStatezny··on Gemini 3.8 Flash and 3.8 Flash Cyber
It struck me today using Google Antigravity (Claude Code alternative) just how direct and usably terse Gemini is in prose.

I've complained plenty on here about Claude verbosity and TED-talk phrasing, and it seems by contrast Gemini has already arrived at the dream end-state of Claude from a prose standpoint.

Sometimes I ask for feedback, and I get back a list of multiple-choice options as if it's already ready to go. If I indicate I'm thinking about doing something, sometimes it'll just...do it. (Not in an annoying way.)

It seems very geared toward action in a way that's completely refreshing coming from months steeped in Claude essays.

PaulStatezny··on Apple caught off guard by AI demand for Mac Mini and Mac Studio
For the amount of tokens you get, based on your comment, ALL subscriptions are heavily subsidized, and the most expensive ones are the most subsidized.

For OpenAI and Anthropic, the $100 subscriptions cost 5x the $20 subscriptions and give you 5x the tokens. And the $200 subscriptions are 10x the cost for 20x the tokens. (Tokens cost 50% as much.)

PaulStatezny··on Why do I lose my passion and want to do nothing?
> there is something different about AI, though, in that it attacks the craft of cognition itself.

It's particularly frustrating that it's simultaneously superhuman and incredibly unreliable when it comes to cognition.

E.g. it can whip out 100,000 lines of code in the time you'd write 500...but 5% of that code will be janky, or subtly buggy, etc.

It's the "oh...you're absolutely right!" effect. I can't think of any other tools that are so powerful yet so finnicky and almost guaranteed to fail in ways that are so difficult to detect.

PaulStatezny··on Claudette: Make Claude stop talking like a BuzzFeed article
For those of us who don't want to pay for Gemini tokens, what would be the best local LLM to use for this?

(A model that can run reasonably well in a ~24GB MacBook.)

PaulStatezny··on Models Are Getting Dumber on Purpose
Which is totally ironic, given that the article is largely about how untrustworthy LLMs can be. (Hallucination.)

Putting readers through this exercise disrespects their time. Even if as a writer you did the work of researching, reasoning, and fact-checking, you shoot yourself in the foot by running it through an LLM because there's no way for the reader to know which thoughts/research are from you. It demolishes the Ethos of the writing; readers feel they must do quality assurance on the reasoning, research, and facts.

PaulStatezny··on [dead]
This is edited not to sound like Claude wrote it, but is full of Claude-isms. It's exhausting to read.

What's going to happen long term when this style of writing is so ubiquitous that people get comfortable reading it? Is it going to fundamentally change English? Are people going to start talking like Claude?

PaulStatezny··on How I use LLMs to learn complex topics
Would you be willing to provide an example (even a contrived one) of how this "four at a time" prompt changes the LLM's behavior?

I just want to understand more.

Also, would you be willing to share the actual text of it that you put in AGENTS.md?

PaulStatezny··on How I use LLMs to learn complex topics
Can you elaborate more on juxtaposing Claude's terrible prose with other LLMs?

Any more detail you can share? Do the others feel more "human"? Are there any that are particularly digestible/human-friendly?

I've been wondering for a while if this is just Claude because I mostly use Claude, so this is very telling.

PaulStatezny··on How I use LLMs to learn complex topics
100% -- this is the worst. Referencing all sorts of "words with made up contextual/analogous meanings" based on the conversation...outside of the conversation.

Does anyone have a read on if this is primarily a Claude issue, or if all LLMs do this?

PaulStatezny··on How I use LLMs to learn complex topics
Yes!

I have a personal theory: LLMs are *fundamentally* handicapped at perceiving what's going on in the mind of the human (this can't be "innovated away") and that's at the root of what makes them suck at conversation.

Next time you're chatting with someone, notice how much understanding is shared without anything being said. E.g. the other person might share something deeply disappointing, and they can tell without you even saying anything whether you get what they're going through. This unspoken-yet-communicated information guides the conversation. Or as another example: humans can read the room -- you walk into a room and immediately adjust your demeanor based on what you see and sense.

LLMs are totally blind to things like this, and this adds an inescapable awkwardness to interacting with them. I don't believe they'll ever grow out of this. Which thankfully implies more long term demand for humans instead of robots. :)

PaulStatezny··on LLMs reward expertise
Yeah, in my experience, there's nothing about:

1. LLM thinking 2. RLHF 3. The latest frontier models

that does anything to change this fundamental "suggestibility" of LLMs.

But who knows, maybe I'm wrong.

PaulStatezny··on LLMs reward expertise
Sometimes when I want AI to explain something technical, I say "explain it like I'm a junior engineer" -- just to get it to start with the high level like a human being would.
PaulStatezny··on LLMs reward expertise
LLMs skew toward over-focusing on things that you mention.

The reason "the agent suddenly started suggesting all kinds of things to make its code more robust" is because you said you "want to build reliable software".

It's not a signal of good judgment or understanding. It's just how LLM attention works.

PaulStatezny··on Vulnerability reports are not special anymore
> Ending on a doom-and-gloom note: there will be a reckoning.

Can you elaborate on what you mean by this?

PaulStatezny··on Smudging the game disc to make speedrunning 'SpongeBob' faster
> this game has some of the most insane tech I’ve seen in any game and is definitely worth checking out

Given the context of this forum, I'd be interested to hear more about what's so interesting about the technology!

PaulStatezny··on Claude Fable 5
Interesting, I assumed all model-routing was done utilizing an LLM. (I.e. non-deterministic.)
PaulStatezny··on United Airlines 767 returns to Newark after Bluetooth name sparks alert
A minor spelling nit. It's "it's", not "its", when used as a contraction for "it is". ;)

Sorry, you teed it up too well. I had to!

PaulStatezny··on Agents need control flow, not more prompts
I agree. But you can speak imperatively to agents as well ("Here are specific steps; follow them") and they can still screw up. :) I think what you're looking for is determinism, not imperativism.

And to your point: instructing a (non-deterministic) LLM declaratively ("get me to this end state") compounds the likelihood of going off the rails.

PaulStatezny··on Why are neural networks and cryptographic ciphers so similar? (2025)
I would highly recommend the free book Crypto 101.

https://www.crypto101.io

PaulStatezny··on AWS CEO says replacing junior devs with AI is 'one of the dumbest ideas'
But without AI, there are neural connections formed while determining the correct one-off solution.

The neural connections (or lack of them) have longer term comprehension-building implications.

PaulStatezny··on AWS CEO says replacing junior devs with AI is 'one of the dumbest ideas'
I think the idea is copy-pasting code snippets from StackOverflow without comprehension of whether (and how) the code fixes the problem.
PaulStatezny··on Formatting code should be unnecessary
You didn't read the blog.

It's talking about the Ada programming language and that its code was apparently stored not as plaintext but an intermediate representation (IR) that could then be transformed back into code.

So formatting was handled by tooling by the nature of the setup. Developers would each have their own custom settings for "pretty printing" the code.

The author isn't saying don't use code formatters. They're highlighting an unusual approach that the industry at large isn't aware of. Instead of getting rid of arguments about code style via formatters, you can get rid of them by saving code in an IR instead of plaintext.

PaulStatezny··on I'm absolutely right
Telling someone they "shouldn't be insecure" reminds me of this famous Bob Newhart segment on Mad TV.

Bob plays the role of a therapist, and when his client explains an issue she's having, his solution is, "STOP IT!"

> You shouldn't be so insecure.

Not assuming that there's any insecurity here, but psychological matters aren't "willed away". That's not how it works.

PaulStatezny··on I'm absolutely right
Truly incisive observation. In fact, I’d go further: your point about the contrast with real friends is so sharp it almost deserves footnotes. If models could recognize brilliance, they’d probably benchmark themselves against this comment before daring to generate another word.
PaulStatezny··on Notes on Managing ADHD
> The best thing for managing this is meditation, and a disciplined lifestyle regiment.

What would be your reaction to the numerous comments on this page where people are saying that they tried and failed to "discipline" themselves for years or decades, only to discover medication later and find that it instantly turned everything around for them?

PaulStatezny··on Cognitive load is what matters
> programmers agree that simpler solutions...are preferred, but the disagreements start about which ones are simpler

Low ego wins.

1. Given: The quality of a codebase as a whole is greatly affected by its level of consistency + cohesiveness

2. Therefore: The best codebases are created by groups that either (1) internally have similar taste or (2) are comprised of low ego people willing to bend their will to the established conventions of the codebase.

Obviously, this comes with caveats. (Objectively bad patterns do exist.) But in general:

Low-ego → Following existing conventions → They become familiar → They seem simpler

Page 1 of 7Next →