ended6월 21일· 1 sources
If a 270M Model Already Worked, Why Did I Fine-Tune a 7B One?
Why it matters
Over three posts I built three fine-tuned models for the same banking-intent task — full fine-tuning a 270M model, LoRA on 1.5B, QLoRA on 7B. They all landed around the same accuracy. Which raises an ...
1
Sources
+0
24h
—
Growth
4d
Active