I’m not sure about the Speed chart. I would expect gpt-4-turbo to be faster than plain gpt-4.
OpenAI are doing a ton of load balancing, presumably constantly tweaking batch sizes to try to optmize across all their workloads.
You can test the GPT-4 vs GPT-4 Turbo on Playground to intuitively confirm that the speeds are similar.
Could the data have been collected when the system is under different loads?
Clearly OpenAI is throttling their API to save costs and get more out of fewer GPUs.