On the other hand, the more mainstream a programming language, the higher proportion of the training data is going to be terrible code.
I think there's an optimal ratio somewhere
I think there's an optimal ratio somewhere
These days I moved up the ladder of abstraction, so I don't really look; the main criteria I have is how the LLM gets things done.