You really only need one example to offset this anecdote, so here you are: https://mobile.twitter.com/mitsuhiko/status/1410886329924194...
Copyright violations are a genuine concern from the outputted code, GitHub themselves have admitted it may emit raw training data rarely.