Pretty much. And, interestingly, even though it is a little bit slower than most other models I can run here it is quite a bit more efficient with the number of tokens used so on longer runs it often comes out ahead in terms of wall clock time. YMMV of course this is for pretty specific workloads and there is a good chance that on other workloads it would perform worse. But so far it has been a pretty good model for us.