in my case with a 6900xt:
1. sudo pacman -S ollama-rocm
2. ollama serve
3. ollama run deepseek-r1:32b
1. sudo pacman -S ollama-rocm
2. ollama serve
3. ollama run deepseek-r1:32b
I tried running a model larger than ram size and it loads some layers into the gpu but offloads to the cpu also. It's faster than cpu alone for me, but not by a lot.