Wow - 3 bits per element. Wonderful work with the quantization, I'd love to implement this. How would you recommend I proceed?
The NNSE paper has associated code already, but I found setting the sparseness preference parameter was very hit-and-miss, which is why I preferred the explicit sparse-by-percentage measure in my work.