Reinforcement learning towards broadly and persistently beneficial models (alignment.openai.com) 1 points by jawiggins 2mo ago ↗ HN
0 comments
[ 2.7 ms ] story [ 10.0 ms ] threadNo comments yet.