Analyzing OpenAI's Reinforcement Fine-Tuning: Less Data, Better Results (openpipe.ai) 4 points by kcorbitt 1y ago ↗ HN
0 comments
[ 1.7 ms ] story [ 15.5 ms ] threadNo comments yet.