Reinforcement Learning from Human Feedback: When the Math Ain't Enough (evalovernite.substack.com) 1 points by scoresmoke 3y ago ↗ HN
0 comments
[ 5.4 ms ] story [ 6.5 ms ] threadNo comments yet.