Edit: yes. Discussion on hn: https://news.ycombinator.com/item?id=17231593
So far Nvidia does excellent technical work with their hardware + their out of the box CUDA toolkit + support in every major deep learning library. The drawback is that they keep everything closed and they charge a huge premium. AMD has great hardware at good prices but the software side is non-existent and the Linux support goes from bad to worse.
Nvidia is the one having absolute control here and they choose to squeeze the market because they can.
Intel's free software release seems to me to be rather in their favour compared with other vendors.
MKL’s fft isn’t a huge improvement on fftw with avx512, but its blas is ~3x as fast as openblas currently. And before these open source projects caught up, it was by far the best.
And I agree, they’ve been better about it lately. I was comparing Intel then and Nvidia now.
For a free avx512 BLAS, use the current release of BLIS. OpenBLAS recently gained skx (but not knl) gemm support, but I don't know how good it is as I don't have the hardware.
When they realized the mistake and came up with SPIR it was too late.
Also SYSCL is no better than CUDA with their community edition.