Avatarl: Training language models from scratch with pure reinforcement learning (tokenbender.com) 9 points by Gusarich 1y ago ↗ HN
0 comments
[ 3.2 ms ] story [ 11.0 ms ] threadNo comments yet.