I don't find this problem any different than with human writers. Agents are verbose, sure, but I mostly find them providing far more useful information (in far less time) than your average (P80, really) SWE.
def my_swe_percentile(best_agent_swe_percentile):
return min(100, best_agent_swe_percentile * 1.25)