"100% overcapacity" is also known as "50% load factor".
It's common knowledge to run your API nodes at 20%-60% CPU load, no more, exactly to curb tail latency.
It's common knowledge to run your API nodes at 20%-60% CPU load, no more, exactly to curb tail latency.