Loading video...
Video Failed to Load
MLX Swift example can also QLoRA fine-tune Llama 3.2. Here's the 1B fine-tuning on my iPhone 15 Pro at > 150 toks/sec. A this rate only takes a few minutes to learn some decent adapters fully on-device.
91,961 views • 2 years ago •via X (Twitter)
10 Comments

Awni Hannun2 years ago
Example is also open-source here (under MIT license):

GPT.Biz2 years ago
This looks amazing! It’s incredible to see fine-tuning done so quickly on a phone

⚡️Dylan White2 years ago
Whoa, that's impressive! Fine-tuning Llama 3.2 on your iPhone 15 Pro is no joke!

involuntarily incelibate2 years ago
If only iOS would expose more substantial endpoints, you could fine tune “Siri”!

Janaka2 years ago
Impressive

Andres Gomez Sarmiento2 years ago
really cool

Kalin Ovtcharov2 years ago
Ability to run llms locally on-device and fine tune them is an extremely important yet relatively under-appreciated/ignored capability today. I expect that will change as new applications emerge that leverage SLMs. I expect some governments may even ban them in the near future.

Josep M. Ganyet2 years ago
@Scobleizer @xaviviro

Sorayesh Semo2 years ago
Awesome

samuel abraham gomez villafuerte2 years ago
Para andriod cuando
