Finetuning a Reasoning LLM with Supervised or Reinforcement Learning? (discuss.huggingface.co) 2 points by verdverm 2mo ago ↗ HN
0 comments
[ 2.9 ms ] story [ 10.5 ms ] threadNo comments yet.