I tried generating the same test with all 5 models in Qodo Gen.
o1 is very slow - like, you can go get a coffee while it generates a single test (if it doesn’t time out in middle).
o1-mini thought worked really well. It generated a good test and wasn’t noticeably slower than the other models.
My feeling is that o1-mini will end up being more useful for coding than o1, except for maybe some specific instances where you need very deep analysis