Loading video...

Video Failed to Load

Go Home

Llama 4 Scout and Maverick from Meta are now live on GroqCloud™. Day-zero access. Real-time performance. Lowest cost—without compromise. No waiting. No tuning. Just build fast.

154,819 views • 1 year ago •via X (Twitter)

11 Comments

Groq Inc's profile picture
Groq Inc1 year ago

Try it →

PowerBeatsVR's profile picture
PowerBeatsVR1 year ago

Box, dodge, and squat your way through PowerBeatsVR - Now 40% OFF on Meta Quest for a limited time 🔥

Groq Inc's profile picture
Groq Inc1 year ago

Read more about Groq performance, price, & more in our blog →

Jake Dahn's profile picture
Jake Dahn1 year ago

@Meta is maverick actually live? all im seeing is `meta-llama/llama-4-scout-17b-16e-instruct` ... also, are y'all supporting the full context window in both models?

Data's profile picture
Data1 year ago

@Meta Is this full precision?

Nayer ALI MAHOMED's profile picture
Nayer ALI MAHOMED1 year ago

@Meta @GroqInc I’m only seeing access to the Scout model. Is Maverick limited or just not rolled out to everyone yet? 🤔

ZOHEB's profile picture
ZOHEB1 year ago

@Meta Groq’s team works as fast as Groq’s responses

V's profile picture
V1 year ago

@Meta That was fast, just out of the oven

✦ xyz's profile picture
✦ xyz1 year ago

@Meta i dont see maverick.. only scout

Romulus's profile picture
Romulus1 year ago

@Meta holy fuck how fast

𝕃𝕏𝔼's profile picture
𝕃𝕏𝔼1 year ago

@Meta With the native full context windows? Damn this might kill RAG.

Related Videos

Introducing "Building with Llama 4." This short course is created with Meta AI at Meta, and taught by Amit Sangani, Director of Partner Engineering for Meta’s AI team. Meta’s new Llama 4 has added three new models and introduced the Mixture-of-Experts (MoE) architecture to its family of open-weight models, making them more efficient to serve. In this course, you’ll work with two of the three new models introduced in Llama 4. First is Maverick, a 400B parameter model, with 128 experts and 17B active parameters. Second is Scout, a 109B parameter model with 16 experts and 17B active parameters. Maverick and Scout support long context windows of up to a million tokens and 10M tokens, respectively. The latter is enough to support directly inputting even fairly large GitHub repos for analysis! In hands-on lessons, you’ll build apps using Llama 4’s new multimodal capabilities including reasoning across multiple images and image grounding, in which you can identify elements in images. You’ll also use the official Llama API, work with Llama 4’s long-context abilities, and learn about Llama’s newest open-source tools: its prompt optimization tool that automatically improves system prompts and synthetic data kit that generates high-quality datasets for fine-tuning. If you need an open model, Llama is a great option, and the Llama 4 family is an important part of any GenAI developer's toolkit. Through this course, you’ll learn to call Llama 4 via API, use its optimization tools, and build features that span text, images, and large context. Please sign up here:

Andrew Ng

67,846 views • 1 year ago