Video yükleniyor...
Video Yüklenemedi
Full Fine-Tuning vs LoRA by hand ✍️ interactive diagram. Open Here's some T/F questions to test your knowledge: [ ] Growing the batch size grows the number of trainable parameters [ ] Both rows produce an output of the same shape [ ] LoRA's ΔW is a smaller matrix... show more
10,063 görüntüleme • 1 ay önce •via X (Twitter)
2 Yorum

Buswe1 ay önce
Worth adding: B starts at zeros, so at step 0 the LoRA branch contributes nothing and the output matches the base model exactly.

Hadi Saeed1 ay önce
This is 10/10. Everyone explains what LoRA is, almost nobody shows why it's cheap. Dragging the rank and watching the count drop makes it click
