Загрузка видео...
Не удалось загрузить видео
Full Fine-Tuning vs LoRA by hand ✍️ interactive diagram. Open Here's some T/F questions to test your knowledge: [ ] Growing the batch size grows the number of trainable parameters [ ] Both rows produce an output of the same shape [ ] LoRA's ΔW is a smaller matrix... show more
10,069 просмотров • 1 месяц назад •via X (Twitter)
Комментарии: 2

Buswe1 месяц назад
Worth adding: B starts at zeros, so at step 0 the LoRA branch contributes nothing and the output matches the base model exactly.

Hadi Saeed1 месяц назад
This is 10/10. Everyone explains what LoRA is, almost nobody shows why it's cheap. Dragging the rank and watching the count drop makes it click
