I just wish they'd support Pascal :(
Nvidia might have given up support but it doesn't mean vllm have to (llama.cpp didn't).
Nvidia might have given up support but it doesn't mean vllm have to (llama.cpp didn't).
https://en.wikipedia.org/wiki/Pascal_(microarchitecture)
I guess many would like to use stock vLLM to run inference on affordable P100...