正在加载视频...

视频加载失败

Llama3.1 on a raspberry pi Thanks for Llamafile 😍

126,760 次观看 • 2 年前 •via X (Twitter)

10 条评论

Mike Bird (Hiring) 的头像
Mike Bird (Hiring)2 年前

I used the @Raspberry_Pi 5 with the M.2 Hat with the Hailo AI module

Tom Dörr 的头像
Tom Dörr2 年前

@JustineTunney Now do 405B

Mike Bird (Hiring) 的头像
Mike Bird (Hiring)2 年前

@JustineTunney

John T Davies 🇺🇦🇪🇺🌍 的头像
John T Davies 🇺🇦🇪🇺🌍2 年前

@JustineTunney My Raspi5 runs Llama3.1 8B (Q4) at just over 1.8 toks/sec, painful if you're waiting for an answer. Qwen2 1.5B (Q4) gave me a pretty good answer for the same example at over 8 toks/sec. We're getting there!

Mike Bird (Hiring) 的头像
Mike Bird (Hiring)2 年前

@JustineTunney Incremental progress!

Prince Canuma 的头像
Prince Canuma2 年前

@JustineTunney That’s pretty cool! Btw if you have a Mac l, you can stream the same model (4-bit quant) much faster using fastMLX (+100 tokens/s for M3 Max 96GB) You can even connect multiple Pis to the same server and run parallel requests :)

Mike Bird (Hiring) 的头像
Mike Bird (Hiring)2 年前

@JustineTunney Definitely will experiment! Thanks Prince!

AshutoshShrivastava 的头像
AshutoshShrivastava2 年前

@JustineTunney Coolest thing on internet today.

Mike Bird (Hiring) 的头像
Mike Bird (Hiring)2 年前

@JustineTunney 🫡

Chubby♨️ 的头像
Chubby♨️2 年前

@JustineTunney I love to see that SLM is getting more and more great possibilities. To think that an SLM as excellent as Llama 3.1 8b is already running on a Raspberry pi, I can't imagine where we'll be in a year's time. Great work

相关视频