Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewritingarxiv.org2 points·djhu9··0 commentsOpen articleSaveView on HN