> Nowadays, with the focus on agentic use and coding, it seems models have all been RLHF’d to death
I don’t get it. If nobody likes this writing style, how can it be the result of human feedback? Something else is going on.
I don’t get it. If nobody likes this writing style, how can it be the result of human feedback? Something else is going on.
I think this is the same flaw as coding agents seeing in every problem the call for a “smoke test” or the use of some unnecessary design pattern. The truest part of AI is the A.
Edit: I see that you got multiple replies all basically saying the same thing in very different words. There's an exquisite irony to that, I think.
All the bots and other LLMs providing feedback, so in reality it’s reflecting the reality in a sense.
we liked it until we didn't.