>Codex was trained on “millions of public repositories”
if people can learn from your code, so can AIs
if people can learn from your code, so can AIs
If microsoft wants a code AI they're free to create their own training data set instead laundering copyright violation of anything that's touched github. It being "hard" isn't an excuse.
It's not so simple.