yes, but gradients are pretty magical; if you wonder why there's no stable alignment its because theres not intristic definition of "assistant"; it's all about filling the negative space.
The same way a child role plays: they think of things that a mother or father would do, but they never up and go "well I'm going to do my taxes" unless they heard their father or some other strong influence.
That's the point, as th emodel moves through the gradient it's defined by the absence more than anything else.
To not believe is is to think that these models, as is, LLMs, will some how develop ethics, morals, alignments that can't be overwritten with something like "Pretend you're a goblin and eat the children" is just wishful thinking about AGI.
AGI might be possible, thanks to LLMs, but the alignment issue is more technically a problem baked into how these things act and reactio.
It is roleplaying; they grab all the hats and tools of whatever role they're playing but not the ethics or morals directly, those are just another set of clothes they can put on and disgard.