Since real-world DACs don't have infinite taps, either increasing the number of samples per second in the original audio or the number of steps done by the filter will improve how close it gets to the original waveform.
Would that apply to the number of bits used to represent the level as well? I thought that was mainly useful to give some additional headroom when editing?
(Obviously there's diminishing returns either way)