Good question - one followup question there is value for who?
If it is to train the LLM that is labeling, then I agree.
If it is to train a smaller downstream model (e.g. finetune a pretrained BERT model) then the value is as good as coming from any human annotator and only a function of label quality