Watching a language model think before it speaks
blog.nathanlangley.dev
blog.nathanlangley.dev
Such as “would you take over the world?”
Then seeing a large enthusiastic “ABSOLUTELY” when the output says “I’m just a wee helpful little AI model…”
This article also reads with many LLM-isms, so I can’t tell if a human actually produced it.