Has anyone tried RLHF on the GPT models when fine-tuning? Would this be useful?2 points·jmiran15··0 commentsSaveView on HN