jobhadel··on Emerging Reasoning with Reinforcement LearningChain of thought breaks down a problem into smaller chunks which is easier to solve for a model than trying to find solution directly for larger problem