You cannot compare GPT-4o and o*(-mini) because GPT-4o is not a reasoning model.
Sure you can.
Even though one is more appropriate for certain tasks than the other.
I do think it is a good metaphor for how all this shakes out though in time.
Ends before means.
If 4o answered better than o3, would you still use 03 for your task just because you were told it can "reason"?