46,301 karma · joined June 2, 2014
email: breen85 [at] gmail [dot] com
Also I just had to upgrade my Vercel plan due to traffic, which is very welcome and appreciated, but if anyone wants to donate, that would be really helpful! There's a Stripe link at the site: https://book-prize-index.vercel.app
And any other ideas that HN readers have for awards to add would be welcome. Currently it's probably too history-slanted since I'm a historian and knew those awards better.
I was thinking about the immortal Twin Peaks line "there's a FISH... in the PERCOLATOR" when I wrote the trout one.
The latter book has a Wikipedia page with some more info - was surprised to see Hacking not mentioned here since the featured article is partly based on his work: https://en.wikipedia.org/wiki/The_Taming_of_Chance
I find this fascinating because it literally just happened in the past few months. Up until ~summer of 2025, the SVG these models made was consistently buggy and crude. By December of 2026, I was able to get results like this from Opus 4.5 (Henry James: the RPG, made almost entirely with SVG): https://the-ambassadors.vercel.app
And now it looks like Gemini 3.1 Pro has vaulted past it.
However, I think there is also something qualitatively different about how work is done in these two domains.
Example: refactoring a codebase is not really analogous to revising a nonfiction book, even though they both involve rewriting of a sort. Even before AI, the former used far more tooling and automated processes. There is, e.g., no ESLint for prose which can tell you which sentences are going to fail to "compile" (i.e., fail to make sense to a reader).
The special taste or skillset of a programmer seems to me to involve systems thinking and tool use in a different way than the special taste of a writer, which is more about transmuting personal life experiences and tacit knowledge into words, even if tools (word processor) and systems (editors, informants, primary sources) are used along the way.
Sort of half formed ideas here but I find this a really rich vein of thought to work through. And one of the points of my post is that writing is about thinking in public and with a readership. Many thanks for helping me do that.
I don't have a good answer to your question, but I do think it might be comparable, yes. If you had good taste about what to get Opus 4.6 to write, and kept iterating on it in a way that exposes the results to public view, I think you'd definitely develop a more fine grained sense of the epistemological perspective of a writer. But you wouldn't be one any more than I'm a software developer just because I've had Claude Code make a lot of GitHub commits lately (if anyone's interested: https://github.com/benjaminbreen).
Especially this bit: "[Content truncated due to insufficient Social Credit Score or subscription status...]"
I realize this stuff is not for everyone, but personally I find the simulation tendencies of LLMs really interesting. It is just about the only truly novel thing about them. My mental model for LLMs is increasingly "improv comedy." They are good at riffing on things and making odd connections. Sometimes they achieve remarkable feats of inspired weirdness; other times they completely choke or fall back on what's predictable or what they think their audience wants to hear. And they are best if not taken entirely seriously.
Enjoy!
I also asked Opus 4.5 to make a "1994 style readme page" for the GitHub: https://github.com/benjaminbreen/HyperCardHackerNews
I'm going to go ask Claude Code to create a functional HyperCard stack version of HN from 1994 now...
Edit: just got a working version of HyperCardHackerNews, will deploy to Vercel and post shortly...
"The immigrant, on arriving, found himself a stranger, in a strange land, far from friends. Time pressed, for the little means that could be realized from the sale of what was left of the outfit would not support a man long at California prices. Many became discouraged. Others would take off their coats and look for a job, no matter what it might be. These succeeded as a rule. There were many young men who had studied professions before they went to California, and who had never done a day's manual labor in their lives, who took in the situation at once and went to work to make a start at anything they could get to do. Some supplied carpenters and masons with material—carrying plank, brick, or mortar, as the case might be; others drove stages, drays, or baggage wagons, until they could do better. More became discouraged early and spent their time looking up people who would 'treat,' or lounging about restaurants and gambling houses where free lunches were furnished daily."
Personally I think it absolutely will lead to major changes in historical research. The transcription and translation abilities of transformer models alone are already leading to significant changes and advances. For instance, I'm working on a post about new transformer based OCR tools like Leo that are geared specifically for historical research and led by historians (https://www.tryleo.ai - I'm not involved in the project, just an interested observer).
IMO AI tools will definitely still be used by a minority of historians in a 5-10 year horizon. Historical research is not like some STEM fields where there is a lab-base culture oriented around adopting new tech and finding applications quickly. It's a lot more of a solo, idiosyncratic process of personal research and that is partly why I like it, but it also means that uptake of new tools is much slower. That said, historians do use technology and digital tools all the time and are not inherently adverse to it. It's interesting reading history books from the 1970s, like the works of Lawrence Stone (https://en.wikipedia.org/wiki/Lawrence_Stone) and seeing the footnotes about how the data was encoded in punchcards and analyzed by mainframes. I expect we will be seeing history books by the end of the 2020s that use custom data analytics and tagging tools developed by the historians themselves using vibe coding.
Thanks for the question, will be writing more about this. Feel free to get in touch any time.
Some fun things I've been experimenting with is 1) injecting primary sources from a given time and place into the LLMs contex to further ground it in "reality" and 2) asking the LLM to try to simulate the actual historical language of the era - i.e. a toggle button to switch to medieval French. Gemini flash lite, the only economical model for this sort of thing, is not great at this yet but in a year or so I think it will be a fascinating history and language learning tool.
Have been meaning to write this project up for HN but if anyone wants to try a very early version of it, it's here - you can modify the url to pick a specific year and region or just do the base url for a fully random spawn, i.e. here is Europe in 1348: https://historysimulator.vercel.app/1348/europe
Oxford and Cambridge have a "tutorial" system that is a lot closer to what I would choose in an ideal world. You write an essay at home, over the course of a week, but then you have to read it to your professor, one on one, and they interrupt you as you go, asking clarifying questions, giving suggestions, etc. (This at least is how it worked for history tutorials when I was a visiting student at an Oxford college back in 2004-5 - not sure if it's still like that). It was by far the best education I ever had because you could get realtime expert feedback on your writing in an iterative process. And it is basically AI proof, because the moment they start getting quizzed on their thinking behind a sentence or claim in an essay, anyone who used ChatGPT to write it for them will be outed.
https://aistudio.google.com/app/prompts?state=%7B%22ids%22:%...
The unsolved issue is scale. 5-10 minute Q&As work well, but are not really doable in a 120 student class like the one I'll be teaching in the fall, let alone the 300-400 student classes some colleagues have.
As an experiment I just asked it to "recreate the early RPG game Pedit5 (https://en.wikipedia.org/wiki/Pedit5), but make it better, with a 1970s terminal aesthetic and use Imagen to dynamically generate relevant game artwork" and it did in fact make a playable, rogue-type RPG, but it has been stuck on "loading art" for the past minute as I try to do battle with a giant bat.
This kind of thing is going to be interesting for teaching. It will be a whole new category of assignment - "design a playable, interactive simulation of the 17th century spice trade, and explain your design choices in detail. Cite 6 relevant secondary sources" and that sort of thing. Ethan Mollick has been doing these types of experiments with LLMs for some time now and I think it's an underrated aspect of what they can be used for. I.e., no one is going to want to actually pay for or play a production version of my Gemini-made copy of Pedit5, but it opens up a new modality for student assignments, prototyping, and learning.
Doesn't do anything for the problem of AI-assisted cheating, which is still kind of a disaster for educators, but the possibilities for genuinely new types of assignments are at least now starting to come into focus.
My read is that most likely, it was recorded on an old school reel-to-reel tape recorder. It's entirely possible that the tapes are still sitting on a shelf somewhere in Argentina, though the chances of actually tracking them down are pretty low. I worked with some reel-to-reel tapes that Alan Ginsberg made (now held at Stanford) in the mid-60s (including one where he is talking to Bob Dylan!) and they held up pretty well. Had to use audio editing software to remove tape hiss, but they were not as badly preserved as I expected.