Why not spend a few tokens or so on making a frontier LLM build the classifier from scratch? Then you also know exactly what the classifier is capable of, and whether it’s easily learnable/informative data.
Why would you be confident in feeding garbage to a “cheap and fast” classifier with unknown domain-specific performance?
I know we kind assume omniscience for frontier models, but at this point the evidence is kind of out there.