Loading video...

Video Failed to Load

Go Home

Llama 3 was *still* learning when Meta stopped training it. They only stopped because they decided they needed the GPUs to start testing for Llama 4. AI scaling laws are insane.

414,180 views • 2 years ago •via X (Twitter)

11 Comments

Mckay Wrigley's profile picture
Mckay Wrigley2 years ago

“The models just want to learn.” - Ilya

PowerBeatsVR's profile picture
PowerBeatsVR1 year ago

Box, dodge, and squat your way through PowerBeatsVR - Now 40% OFF on Meta Quest for a limited time 🔥

Justin Halford's profile picture
Justin Halford2 years ago

Even crazier is that they only trained with 16k H100s despite having ordered 600k. So only 2.6% of the compute from that order was used to create a GPT-4 level model. Wild times ahead.

Mckay Wrigley's profile picture
Mckay Wrigley2 years ago

I’m only halfway through this so far, but Zuck even says he thinks energy will be a bottleneck before compute. Wild!

Magic Carpet 🇺🇸's profile picture
Magic Carpet 🇺🇸2 years ago

Almost all LLMs are undertrained relative to size

miru's profile picture
miru2 years ago

@HlibIvanov source of the clip

ASI - Tech Gone Wild 🤖❤️‍🔥⚔️'s profile picture
ASI - Tech Gone Wild 🤖❤️‍🔥⚔️2 years ago

Very Very promising for FSD and Teslas Fleet advantage

λy's profile picture
λy2 years ago

llama 3 still *is* learning there’s a checkpoint on the biggest 400B model today, but they’re still training it

Luci Pars's profile picture
Luci Pars2 years ago

AGI 💖 Mark 🤣

Josh Whiton's profile picture
Josh Whiton2 years ago

“There are these meta reasoning questions…”

EMAD's profile picture
EMAD2 years ago

So maybe scale actually is all you need…?

Related Videos