It's good but (at least on openrouter) it's got an annoyingly tight output token limit. So if it does get stuck in a reasoning pit, it won't work its way out of it in time.
It's replaced the Kimi models for me though.
It's replaced the Kimi models for me though.
Availability and rate limiting can be an issue, though. I’ve found that constraining the providers works well to solve that though.