I was completely surprised that the reasoning comes from within the model. When using gpt-o1 I thought it's actually some optimized multi-prompt chain, hidden behind an API endpoint.
Something like: collect some thoughts about this input; review the thoughts you created; create more thoughts if needed or provide a final answer; ...