Try using a coding agent to write an efficient GPU kernel. I guess they might get good at it soon, but they definitely aren't there yet.
FWIW, this talk[1] from NVIDIA/Meta from March claims that coding agents can often write correct implementations of of CUDA kernels, but that they're usually dog slow, like 100x slower than a kernel optimized by a skilled human.
[1] https://www.nvidia.com/en-us/on-demand/session/gtc26-s81653/
I suspect when people said AI wrote non performant CUDA kernels it was beginning-mid last year and it's definitely vastly improved since back then. And the agent's ability to iteratively improve really impressed me.