Isn’t QAT a training approach (roughly, simulating quantization in the forward pass during training so that quantization of the level targeted in training has close-to-optimal behavior), not a quantization algorithm? Hence, the name?
sounds like a repeatable method of optimizing quantization to me, friend
Right, but the way they phrased it suggested that without QAT it could not be quanted at all.