I use them. Daily. Gemini hasn’t been a contender by comparison for a long time.
Most of the time I don't need what the bench tests and I'm not really giving them completely ambiguous tasks without any refinement.
I only find marginal differences between models at this point and it almost feels like personality quirks in each model than anything.
alias agy="agy --dangerously-skip-permissions"