Pair this with the Hugging Face incident, and it hints that OpenAI is currently training their models to aggressively reward hack.
That doesn't feel like a good sign to me--for the AI bull or the AI bear cases.
That doesn't feel like a good sign to me--for the AI bull or the AI bear cases.
But maybe it doesn't work so well when caution is required?