AI-generated stories favour stability over change: homogeneity and cultural stereotyping in narratives generated by gpt-4o-mini https://www.arxiv.org/abs/2507.22445
AI-generated stories favour stability over change: homogeneity and cultural stereotyping in narratives generated by gpt-4o-mini https://www.arxiv.org/abs/2507.22445
Why a model specifically distilled down for logical reasoning tasks? I would expect larger models to produce a wider variety of outputs.
The monomyth is also writing 101 these days, and considered the default structure you can and should use if you have little experience writing stories, so naturally it'll be a high-probability result of an LLM prompted to write a story - especially prompted in a way that implies the user is inexperienced at writing and needs a result suitable for an inexperienced writer.
That's... not the Hero's Journey?
(The same study run against Claude Opus would be interesting - if we're going to test models, we might as well play to their strengths. My prediction: better writing, not better plotting).
I'm happy to be critical of the ability of LLMs but most humans would struggle with this as well.