3 comments

[ 2.8 ms ] story [ 54.5 ms ] thread
LoRA AWS swarms are 60x 1 GPU throughput for fine-tuning?? Kind of a big find
How is the model performance? I heard lora on LLM generally get inferior results comparing to no LORA finetuning
Generally good as long as you target all layers.