Your configuration is broken or wrong. What are you using? Hopefully not llama.cpp?
I’ve sweeped concurrency across many models and many different kinds of hardware, and the only times I saw similar results to you were when I didn’t configure it correctly.