Now I dismiss it every time and the quality is more consistent.
Complete adhoc and personal experience but something I've observed, wouldn't be surprised if they nerfed on a per session basis
Now I dismiss it every time and the quality is more consistent.
Complete adhoc and personal experience but something I've observed, wouldn't be surprised if they nerfed on a per session basis
Random performance is random, your brain will jump through hurdles to fit patterns where there aren't any.
In fact, since they have some rule based system (if bioenegineering or security, route to degraded model) it would be almost trivial to add 'user has filed feedback' to it.
Not saying this is what happens, but just that it's not as insane as it sounds.
And it's also the kind of solution a misaligned agent would implement:
Make sure users are happy about their experience but also optimize resources.
The obvious solution is to route users based on feedback.
Using feedback on your sessions makes that session, and presumably all attached data, fair game for training.
They explicitly ask, after the rating, whether they can look at your chat. You can just say "no" to that, if you believe that they are following the rules they say they are.
Your original statement of "using feedback gives them permission to train" is just plain false.
And strangely, expressing frustration multiple times in a row would reliably trigger a feedback popup as well.
That said, given the propensity for mature code bases to have "fuck" in commit messages/comments and those are typically of higher quality, I curse up a storm when the clankers make mistakes, if only to try put more quality-code valence into context. https://news.ycombinator.com/item?id=36584464