Video wird geladen...
Video konnte nicht geladen werden
MLX Swift example can also QLoRA fine-tune Llama 3.2. Here's the 1B fine-tuning on my iPhone 15 Pro at > 150 toks/sec. A this rate only takes a few minutes to learn some decent adapters fully on-device.
91,961 Aufrufe • vor 2 Jahren •via X (Twitter)
10 Kommentare

Awni Hannunvor 2 Jahren
Example is also open-source here (under MIT license):

GPT.Bizvor 2 Jahren
This looks amazing! It’s incredible to see fine-tuning done so quickly on a phone

⚡️Dylan Whitevor 2 Jahren
Whoa, that's impressive! Fine-tuning Llama 3.2 on your iPhone 15 Pro is no joke!

involuntarily incelibatevor 2 Jahren
If only iOS would expose more substantial endpoints, you could fine tune “Siri”!

Janakavor 2 Jahren
Impressive

Andres Gomez Sarmientovor 2 Jahren
really cool

Kalin Ovtcharovvor 2 Jahren
Ability to run llms locally on-device and fine tune them is an extremely important yet relatively under-appreciated/ignored capability today. I expect that will change as new applications emerge that leverage SLMs. I expect some governments may even ban them in the near future.

Josep M. Ganyetvor 2 Jahren
@Scobleizer @xaviviro

Sorayesh Semovor 2 Jahren
Awesome

samuel abraham gomez villafuertevor 2 Jahren
Para andriod cuando
