And now there are updated numbers in the graphs (labeled "gunicorn-3w") for gunicorn, based on running it the way it's meant to be run (on a socket with more than one worker). The results:
* Concurrent requests more than doubled
* Response times dropped 75%
* Error rate dropped 75%
It's amazing what a difference it makes to actually run something the way it was designed to run...