ParentFull threadgenpfault·Getting ~150 tok/s on an empty context with a 24 GB 7900XTX via llama.cpp's Vukan backend.View on HN