New deepseek paper: Natively Trainable Sparse Attention mechanismtwitter.com5 points·redlock··1 commentOpen articleSaveView on HN