Re-quantizing a local LLM 14x faster by skipping the tensors that didn't change | Hacker News Reader