Distillation a pretty well documented technique that actually pre-dates LLMs https://arxiv.org/pdf/1503.02531
Here is a project that guides you through it if you want to prove to yourself that it works https://github.com/arcee-ai/DistillKit
Here is a project that guides you through it if you want to prove to yourself that it works https://github.com/arcee-ai/DistillKit
GPU kernel optimization is just the kind of well-bounded problem with clear success criteria that AI loves.
It's about evidence this is an active force in competition in LLMs.