Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Next level: QLoRA fine-tuning 4-bit Llama 3 8B on iPhone 15 pro. Incoming (Q)LoRA MLX Swift example by David Koski: works with lot's of models (Mistral, Gemma, Phi-2, etc)

581,777 görüntüleme • 2 yıl önce •via X (Twitter)

10 Yorum

Awni Hannun profil fotoğrafı
Awni Hannun2 yıl önce

Stats: - 4 LoRA layers, batch size 1 - runs at about 45-50 toks/sec

Cavit Erginsoy profil fotoğrafı
Cavit Erginsoy2 yıl önce

Cool, but why 🤗

Ryan Ziolko 🚀 profil fotoğrafı
Ryan Ziolko 🚀2 yıl önce

Oh no, Wi-Fi stayed on in airplane mode (fear not, I believe you). Pretty wild to see fine-tuning on a phone. 💪

Daniel Merja profil fotoğrafı
Daniel Merja2 yıl önce

Cluster of IPhones next

SM profil fotoğrafı
SM2 yıl önce

My phones about to become a heater

David profil fotoğrafı
David2 yıl önce

Please forgive the ignorance on my part, but what is it being fine tuned *for* or *on*?

Awni Hannun profil fotoğrafı
Awni Hannun2 yıl önce

In this case wikiSQL. But it’s easy to change the data

Prem Kumar Aparanji 🤖👶🏼🐘 profil fotoğrafı
Prem Kumar Aparanji 🤖👶🏼🐘2 yıl önce

Mindblowing 🤯

Maziyar PANAHI profil fotoğrafı
Maziyar PANAHI2 yıl önce

you gotta be kidding! damn! 👏🏼

Will Townsend profil fotoğrafı
Will Townsend2 yıl önce

Lots of progress for local first LLMs being made, and I'm all here for it!

Benzer Videolar

Introducing "Building with Llama 4." This short course is created with Meta AI at Meta, and taught by Amit Sangani, Director of Partner Engineering for Meta’s AI team. Meta’s new Llama 4 has added three new models and introduced the Mixture-of-Experts (MoE) architecture to its family of open-weight models, making them more efficient to serve. In this course, you’ll work with two of the three new models introduced in Llama 4. First is Maverick, a 400B parameter model, with 128 experts and 17B active parameters. Second is Scout, a 109B parameter model with 16 experts and 17B active parameters. Maverick and Scout support long context windows of up to a million tokens and 10M tokens, respectively. The latter is enough to support directly inputting even fairly large GitHub repos for analysis! In hands-on lessons, you’ll build apps using Llama 4’s new multimodal capabilities including reasoning across multiple images and image grounding, in which you can identify elements in images. You’ll also use the official Llama API, work with Llama 4’s long-context abilities, and learn about Llama’s newest open-source tools: its prompt optimization tool that automatically improves system prompts and synthetic data kit that generates high-quality datasets for fine-tuning. If you need an open model, Llama is a great option, and the Llama 4 family is an important part of any GenAI developer's toolkit. Through this course, you’ll learn to call Llama 4 via API, use its optimization tools, and build features that span text, images, and large context. Please sign up here:

Andrew Ng

67,846 görüntüleme • 1 yıl önce