Optimizing Distributed Training on Frontier for Large Language Modelsarxiv.org2 points·arcanus··0 commentsOpen articleSaveView on HN