Загрузка видео...

Не удалось загрузить видео

На главную

MLX Swift example can also QLoRA fine-tune Llama 3.2. Here's the 1B fine-tuning on my iPhone 15 Pro at > 150 toks/sec. A this rate only takes a few minutes to learn some decent adapters fully on-device.

91,961 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 10

Фото профиля Awni Hannun
Awni Hannun2 лет назад

Example is also open-source here (under MIT license):

Фото профиля GPT.Biz
GPT.Biz2 лет назад

This looks amazing! It’s incredible to see fine-tuning done so quickly on a phone

Фото профиля ⚡️Dylan White
⚡️Dylan White2 лет назад

Whoa, that's impressive! Fine-tuning Llama 3.2 on your iPhone 15 Pro is no joke!

Фото профиля involuntarily incelibate
involuntarily incelibate2 лет назад

If only iOS would expose more substantial endpoints, you could fine tune “Siri”!

Фото профиля Janaka
Janaka2 лет назад

Impressive

Фото профиля Andres Gomez Sarmiento
Andres Gomez Sarmiento2 лет назад

really cool

Фото профиля Kalin Ovtcharov
Kalin Ovtcharov2 лет назад

Ability to run llms locally on-device and fine tune them is an extremely important yet relatively under-appreciated/ignored capability today. I expect that will change as new applications emerge that leverage SLMs. I expect some governments may even ban them in the near future.

Фото профиля Josep M. Ganyet
Josep M. Ganyet2 лет назад

@Scobleizer @xaviviro

Фото профиля Sorayesh Semo
Sorayesh Semo2 лет назад

Awesome

Фото профиля samuel abraham gomez villafuerte
samuel abraham gomez villafuerte2 лет назад

Para andriod cuando

Похожие видео

"Introducing Multimodal Llama 3.2": As promised two weeks ago, here's the short course on Meta's latest open model! This short course is created with Meta and taught by Amit Sangani, Director of AI Partner Engineering at Meta. Meta’s Llama family of models is leading the way in open models, allowing anyone to download, customize, fine-tune, or build new applications on top of them. Learn about the vision capabilities of the Llama 3.2, and use it for image classification, prompting, tokenization, tool-calling. You'll also learn about the open-source Llama stack, which gives building blocks for many different stages of the LLM application life cycle. In detail, you’ll: - Learn what are the features of Meta's four newest models, and when to use which Llama model. - Learn best practices for multimodal prompting, with applications to advanced image reasoning, illustrated by many examples: Understanding errors on a car dashboard, adding up the total of photographed restaurant receipts, grading written math homework. - Use different roles—system, user, assistant, ipython—in the Llama 3.1 and 3.2 models and the prompt format that identifies those roles. - Understand how Llama uses the tiktoken tokenizer, and how it has expanded to a 128k vocabulary size that improves encoding efficiency and multilingual support. - Learn how to prompt Llama to call built-in and custom tools (functions) with examples for web search and solving math equations. - Learn about Llama Stack, a standardized interface for common toolchain components like fine-tuning or synthetic data generation, useful for building agentic applications. By the end of this course, you’ll be equipped to build out new applications with the new Llama 3.2. Thank you to Ahmad Al-Dahle, Amit Sangani, and the whole AI at Meta team AI at Meta for all the hard work on Llama 3.2 — we’re excited to make these open models even more accessible to more developers with this new course! Please sign up here!

Andrew Ng

131,846 просмотров • 2 лет назад