Show HN: Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO) (github.com) 21 points by popopanda 1mo ago ↗ HN
1 comment
[ 0.24 ms ] story [ 26.4 ms ] thread