正在加载视频...

视频加载失败

Llama 3 was *still* learning when Meta stopped training it. They only stopped because they decided they needed the GPUs to start testing for Llama 4. AI scaling laws are insane.

414,180 次观看 • 2 年前 •via X (Twitter)

11 条评论

Mckay Wrigley 的头像
Mckay Wrigley2 年前

“The models just want to learn.” - Ilya

PowerBeatsVR 的头像
PowerBeatsVR1 年前

Box, dodge, and squat your way through PowerBeatsVR - Now 40% OFF on Meta Quest for a limited time 🔥

Justin Halford 的头像
Justin Halford2 年前

Even crazier is that they only trained with 16k H100s despite having ordered 600k. So only 2.6% of the compute from that order was used to create a GPT-4 level model. Wild times ahead.

Mckay Wrigley 的头像
Mckay Wrigley2 年前

I’m only halfway through this so far, but Zuck even says he thinks energy will be a bottleneck before compute. Wild!

Magic Carpet 🇺🇸 的头像
Magic Carpet 🇺🇸2 年前

Almost all LLMs are undertrained relative to size

miru 的头像
miru2 年前

@HlibIvanov source of the clip

ASI - Tech Gone Wild 🤖❤️‍🔥⚔️ 的头像
ASI - Tech Gone Wild 🤖❤️‍🔥⚔️2 年前

Very Very promising for FSD and Teslas Fleet advantage

λy 的头像
λy2 年前

llama 3 still *is* learning there’s a checkpoint on the biggest 400B model today, but they’re still training it

Luci Pars 的头像
Luci Pars2 年前

AGI 💖 Mark 🤣

Josh Whiton 的头像
Josh Whiton2 年前

“There are these meta reasoning questions…”

EMAD 的头像
EMAD2 年前

So maybe scale actually is all you need…?

相关视频