I haven't spun up thinkingcap yet but I'm aware of it and am intending to try it out soon. How did you find it?
I think it does use fewer tokens while reasoning, which is potentially useful. I need to do more testing, because any performance advantage over the 27B is useful for me on an M1 Max.