Explaining Reinforcement Learning with Human Feedback (RLHF) (surgehq.ai) 11 points by echen 3y ago ↗ HN
0 comments
[ 2.3 ms ] story [ 21.8 ms ] threadNo comments yet.