OpenAI will simply set up a classifier to detect if the client is livenerf, and selectively not nerf those requests.
Open models are the endgame.
Open models are the endgame.
The classifier is a model; it examines the actual prompt.
They already do this for the safety "guardrails".