I try really hard to prompt it to replace it with something else. It acknowledges and agree to it. Did it maybe once or twice, then reverted back to the old “As a AI model”
IIRC, I was trying to see if it could replace by “as a LLM”.
IIRC, I was trying to see if it could replace by “as a LLM”.
I try to make it play a games with me and start the prompt differently until a specific keyword was entered… it kinda worked. Kinda being key.
Like, I haven’t heard about a way they could actually implement filters this powerful “inside” the model, it feels like it’s probably a less elegant system than we’d imagine.
They’ve probably done it strongly enough that it can’t really not do it, maybe on purpose to prevent misuse