This was my reaction too. How do you “prove” my gradient is valid?
Well perhaps one way is you could have another LLM take a look at the data you are submitting and have it predict p(useful|not useful) , and create an incentive for users to generate authentic data