I've used Astra, Fable 5.1, Sol 5.6, Opus 5. They're definitely making progress and the proof is in the pudding - I reach for the most advanced model as much as I can. But I wouldn't say their capabilities have increased dramatically. I use it for both coding and non-coding workloads and I don't feel I can do anything dramatically different. Again at the end of the day this is a debate on semantics because we can have very different definitions of "dramatic improvement".
(I'm not factoring the benchmarks into the discussion, because I've never quite cared about them)