Video wird geladen...
Video konnte nicht geladen werden
Full Fine-Tuning vs LoRA by hand ✍️ interactive diagram. Open Here's some T/F questions to test your knowledge: [ ] Growing the batch size grows the number of trainable parameters [ ] Both rows produce an output of the same shape [ ] LoRA's ΔW is a smaller matrix... show more
10,063 Aufrufe • vor 1 Monat •via X (Twitter)
2 Kommentare

Buswevor 1 Monat
Worth adding: B starts at zeros, so at step 0 the LoRA branch contributes nothing and the output matches the base model exactly.

Hadi Saeedvor 1 Monat
This is 10/10. Everyone explains what LoRA is, almost nobody shows why it's cheap. Dragging the rank and watching the count drop makes it click
