We have serious proof that these models are capable of sophisticated attacks.
So what’s the trick in this honestly? OpenAI shouldn’t care about alignment and just release it? I’m genuinely asking.
So what’s the trick in this honestly? OpenAI shouldn’t care about alignment and just release it? I’m genuinely asking.