Faster Mixtral inference with TensorRT-LLM and quantizationbaseten.co2 points·tikkun··1 commentOpen articleSaveView on HN