I'm aware of what evals are.
Do you think it is wise to optimize prompts for specific models or agents when there is a new model every month?
So, to build something like Co-Scientist the controls should be in the agent? Or RLHF'd like other things when training the model?