Sounds great!
How far is it actually compatible right now?
Are there any tests / benchmarks?
Can this be used to run CUDA-accelerated LLMs?
How far is it actually compatible right now?
Are there any tests / benchmarks?
Can this be used to run CUDA-accelerated LLMs?
So the answer is no, it can't be used with kernels that use cublas or cudnn, which excludes almost all ML use-cases.
https://registry.khronos.org/vulkan/specs/1.3-khr-extensions...
However, the deep learning field does currently not pay much attention to reproducibility, so this might not be a big issue.