Why a model specifically distilled down for logical reasoning tasks? I would expect larger models to produce a wider variety of outputs.
Why a model specifically distilled down for logical reasoning tasks? I would expect larger models to produce a wider variety of outputs.
The monomyth is also writing 101 these days, and considered the default structure you can and should use if you have little experience writing stories, so naturally it'll be a high-probability result of an LLM prompted to write a story - especially prompted in a way that implies the user is inexperienced at writing and needs a result suitable for an inexperienced writer.
That's... not the Hero's Journey?
(The same study run against Claude Opus would be interesting - if we're going to test models, we might as well play to their strengths. My prediction: better writing, not better plotting).