It's the backend for torch.compile, pytorch eager mode will still use cuBLAS/cuDNN/custom CUDA kernels, not sure what's the usage of torch.compile
consider that at minimum both FB and OAI themselves definitely make heavy use of the Triton backend in PyTorch.