The silicon is just one piece of the puzzle. CUDA and the rest of the software stack is huge advantage for NVIDIA.
That has historically been the experience with Nvidia GPUs on Linux also.
(Replaced "with 30%" with "within 30%")
I'd really like AMD and Apple to start from scratch with a compute-oriented GPU architecture, ideally standardized with Khronos. The NPU/tensor coprocessor architecture has already proven itself to be a bad idea.
And the big players don't necessarily care about the full software stack, they are likely to optimize the hardware for single usage (e.g. inference or specific steps of the training).
ZLUDA has more interest that SyCL and that should say it all right there.