Very cool. Does anyone know exactly how out of luck us AMD folk are? I know there are efforts out there, but I'm kind of hoping for something "as easy as this?"
Actually, I guess that is more common now that AMD CPUs have a GPU built in?
I don't believe that AMD/NVIDIA is a low entropy bit so far as configuration permutations go. Although NVIDIA is far more widespread, AMD has significant market share. The Darwin bit you already facilitate for is probably lower entropy.
I get the boot concern, and the maintenance concern (!!!), but as you say, these models are already quite huge anyway :)
They basically just ship executables for different llama.cpp backends and select the correct one with a python script, which is fine, as the executables are really small.
https://github.com/YellowRoseCx/koboldcpp-rocm
Some other projects support rocm less explicitly, and not as easily.