Because the original data the model is trained on is under copyright you can't include the output in a copyleft codebase.
See DCO. You cant sign off on you having permission to license the code when you stole it.
Many open source contributors like copyleft, but with illegal LLM copyright-washing many people disrespect the terms. https://en.wikipedia.org/wiki/Developer_Certificate_of_Origi...