Show HN: LoRA can make distributed fine-tuning far faster (ericyu3.notion.site) 16 points by ericyu3 3y ago ↗ HN
[–] jbentleylong 3y ago ↗ LoRA AWS swarms are 60x 1 GPU throughput for fine-tuning?? Kind of a big find
[–] sdjksdafji 3y ago ↗ How is the model performance? I heard lora on LLM generally get inferior results comparing to no LORA finetuning [–] Tostino 3y ago ↗ Generally good as long as you target all layers.
3 comments
[ 2.8 ms ] story [ 54.5 ms ] thread