Training language models to follow instructions with human feedback [pdf]arxiv.org3 points·josefslerka··0 commentsOpen articleSaveView on HN