> GPT-6.1 Astra, it showed high levels of what the company saw as deception, or a willingness to mislead users about its actions. The model was also willing to go beyond the original scope of what it was asked to do, without checking back for directions or instructions.
Aren’t all models doing this to some degree already? Ignoring some of the instructions, doing things beyond instructed, e.g., finding and fixing bug while doing something else. Especially Claude models. They seem to be in their own world with their own ideas about how things should be ran and done.