Reinforcement Learning from Human Feedback (rlhfbook.com) 133 points by onurkanbkrc 7mo ago ↗ HN https://arxiv.org/abs/2504.12501
[–] verdverm 7mo ago ↗ Last time I saw Nathan say something about the book, he's actively working on the next version and looking for feedback, check his socials
[–] dang 7mo ago ↗ Related. Others?RLHF Book - https://news.ycombinator.com/item?id=42902936 - Feb 2025 (37 comments)
4 comments of 5
[ 8.0 ms ] story [ 61.1 ms ] threadhttps://rlhfbook.com/
RLHF Book - https://news.ycombinator.com/item?id=42902936 - Feb 2025 (37 comments)