ParentFull threadjohn_alan·aren't the smaller param models all just Qwen/Llama trained on R1 600bn?View on HN