Has anyone tried RLHF on the GPT models when fine-tuning? Would this be useful?

2 points by jmiran15 ↗ HN

0 comments

[ 4.8 ms ] story [ 6.9 ms ] thread

No comments yet.