This just wrapping llama.cpp right?
I’m sorry but I’m pretty tired of projects wrapping x.cpp.
I’ve been developing a Rust + WebGPU ML framework for the past 6 months. I’ve learned quickly how impressive the work by GG is.
It’s early stages but you can check it out here: https://www.ratchet.sh/ https://github.com/FL33TW00D/whisper-turbo