Yeah, I’m similarly not sure how to square away claims of “trained 100% exclusively on synthetic data” with the fairly obvious fact that it’s derived from an open model, almost all of which are trained on a mix of scraped and data and distilled traces.