Full threadthrowaway888abc·Compressing Large Language Models using Low Rank and Low Precision Decompositionhttps://arxiv.org/abs/2405.18886View on HN