IME, unquantised -> FP8 is pretty much lossless. What matters more is having an unquantized KV cache - using an FP8 KV cache can result in a significant drop in quality.
Claude Shannon is rolling in his grave.
https://en.wikipedia.org/wiki/Rate%E2%80%93distortion_theory
Shannon would be pointing out that if you can throw away half the model without apparent degradation, we're nowhere near packing in all the information we could in training. There must be a better arrangement than we've currently got.
So you could end up paying more for unquantised weights, only to get silently hit with a quantised KV cache...