Loading video...
Video Failed to Load
Next level: QLoRA fine-tuning 4-bit Llama 3 8B on iPhone 15 pro. Incoming (Q)LoRA MLX Swift example by David Koski: works with lot's of models (Mistral, Gemma, Phi-2, etc)
581,723 views • 2 years ago •via X (Twitter)
10 Comments

Awni Hannun2 years ago
Stats: - 4 LoRA layers, batch size 1 - runs at about 45-50 toks/sec

Cavit Erginsoy2 years ago
Cool, but why 🤗

Ryan Ziolko 🚀2 years ago
Oh no, Wi-Fi stayed on in airplane mode (fear not, I believe you). Pretty wild to see fine-tuning on a phone. 💪

Daniel Merja2 years ago
Cluster of IPhones next

SM2 years ago
My phones about to become a heater

David2 years ago
Please forgive the ignorance on my part, but what is it being fine tuned *for* or *on*?

Awni Hannun2 years ago
In this case wikiSQL. But it’s easy to change the data

Prem Kumar Aparanji 🤖👶🏼🐘2 years ago
Mindblowing 🤯

Maziyar PANAHI2 years ago
you gotta be kidding! damn! 👏🏼

Will Townsend2 years ago
Lots of progress for local first LLMs being made, and I'm all here for it!
