Fair enough yeah. IMO microbenchmarks are useful for trying to gain intuition and propose the next incremental improvement. In real world measuring is mandatory, if you don't want to risk pessimizing the code
The wild thing about locks in particular is that microbenchmarks will cause you to make „optimizations” that are the opposite of what benefits real world workloads.