But this is even worse, because there is no way that OpenAI "trained" on this data between September 8, 2026, the date Chenglong Ma uploaded the paper to ChatGPT; and October 6, 2026, the date that OpenAI released a paper with "identical proof strategy and specific choices of notation" (Ma). That's one month, that's not the timescale for model training.
So this implies _not_ that OpenAI is training on user input, in the conventional sense of adjusting weights; but rather that they are *straight-up channeling ideas from user input*, and with a very short lag. You would think there would be about a million controls to prevent this.
This is next-level alarming. I would be very interested in knowing whether Chenlong activated the privacy (do not train, etc) options in ChatGPT, and any other details of their setup (which plan, etc). Also, note that "do not train" might be, in a lawyerly sense, considered by OpenAI to be strictly about weights, and not covering "we hoover up your results and regurgitate them".