803 karma · joined June 28, 2024
Personal site https://tonyalicea.dev
Courses and coaching https://dontimitate.dev
It's Tetris but:
- You have to build a tetromino either manually or have AI "generate" it.
- Generated tetrominoes might mutate after falling.
- You can spend time to "review" generated tetrominoes as they fall to see how they will mutate.
The theme is the risk/reward/pressure of agentic code vs manual vs reviewed agentic code in a deadline environment.
Saying "as long as it works and tests pass" suggests that we can test for every possible scenario. We can't. And tests can be flawed on top of it. Which is why an LLM is no more a compiler than a human coder (as you say) is.
1. Decades of generally reliable systems (I trust my calculator because it always says 1+1=2) has trained people to believe what computers say.
2. LLMs "speak" with great confidence, which has, as you say, a psychological effect.
I've been worried about this a lot, I even created an agent skill called do-i-understand that's designed for novice devs (and experienced too, because atrophy) where the LLM asks you questions about the PR you're about to submit. I've found it helps a lot: https://github.com/AnthonyPAlicea/skills/blob/main/skills/do...
One way or another, there will be a skill reckoning.
I switched to https://bento.page/ when it showed up here on HN and that’s been much more accurate I think because HTML and CSS is even more represented in training data.
It extends beyond search as well. I have had multiple incorrect Gmail summaries that, if I had only read them instead of the actual email, would have resulted in financial harm.
You’re suggesting aligning to an inherently and deliberately changeable presentation.
The idea is fine, but worry about what an HTML element “looks like” is not a good reason.
This seems to conflate appearance with semantics. If an element causes a navigation, I make it a link. Whether it looks like a button is irrelevant, that’s CSS.
I always choose one or the other by intended behavior first, and that always works out great.
That said, I like the idea.
I feel like the old Choose Your Own Adventure books were a sort of precursor to programming for a lot of us. Basic if/then/goto.
I’m not saying I’m hand coding, but saying getting good outputs doesn’t take any technique means your definition of good outputs probably isn’t very refined. And saying it hasn’t taken technique for 4 years is just flat wrong.
Opening up the QBasic source code was eye opening for me: that computers weren't the realm of hackers typing incomprehensible text on TV and swapping floppy disks, but something that could start to make sense. My TRS-80 came next, and, for me, the rest was history.
Good memories. I wonder what the memories will be for kids right now in the age of LLMs.
This experiment was a guidance experiment for me with Fable. Heavy spec and planning architecturally and UX, and strictly no JS frameworks or dependencies.
I do these kind of experiments with new models, and Fable has been the most successful so far for capturing the intended UX and feel per spec.
Things where you know how they’re supposed to make you feel (like games from childhood), and getting that experience match is an exercise.
I think it needs tweaking to let you hone in a value better. What are you thinking?
That is, if you can calculate it too easily then the first player always has a huge advantage.