Loading video...

Video Failed to Load

Go Home

Gemini 4 Pro checkpoints have finally started appearing internally a few days ago. This is the first ever output from the model, internally codenamed "argon". It took 2.4 minutes on High thinking effort. It has a 256k token "output limit", compared to 64k in previous Gemini models. I also...

375,337 views • 4 days ago •via X (Twitter)

21 Comments

Justin's profile picture
Justin3 days ago

256k output limit is insane. If they really fix the lazy character of gemini models it might be a cool model!. Output looks okay :)

J A Z I I's profile picture
J A Z I I3 days ago

minimax can do way better then this SIGH

Harshith's profile picture
Harshith3 days ago

bro is it really Gemini 4 pro why output looks like this??

Mathis Dupuy's profile picture
Mathis Dupuy3 days ago

The bridge is poorly placed samed for the cherry blossom next to the tower … this looks not promising

The Hero of KVcache's profile picture
The Hero of KVcache3 days ago

Selling all my goog stock right now

LeMi's profile picture
LeMi3 days ago

It only took 2.4 minutes, that's a pretty fast model

Louis's profile picture
Louis3 days ago

2.4 minutes is quite fast.

LeMi's profile picture
LeMi3 days ago

is it the RSI model that people keep saying?

Joaoneves's profile picture
Joaoneves3 days ago

This is not good, wtf...

Ritwik's profile picture
Ritwik3 days ago

256k output is the detail that jumps out. That changes the shape of long-horizon tool traces more than another headline benchmark would.

Calimanu Loredan's profile picture
Calimanu Loredan3 days ago

How’s the behavior? Had the chance to test that? Gemini models used to do more, or random stuff even unasked

Ritwik's profile picture
Ritwik3 days ago

256K output is the claim that matters if it's a hard generation cap. Context window rumors are cheap until the tokenizer and cache pricing show up.

ethereagle · building's profile picture
ethereagle · building3 days ago

256k output vs 64k is the number that matters. is that max_tokens, or thinking plus the answer in one budget? High at 2.4 min only helps if reasoning doesn't eat it

lewington's profile picture
lewington3 days ago

256k output is gonna be really nice for coding agents. you can just give them bigger jobs and let them keep going without having to break everything up so much

VDN's profile picture
VDN3 days ago

Amazing news ! Looking forward to it ! I am pretty sure it will be a revolution, because 3.1 pro has aged very well and still is very relevant for many things. Keep up the good work !

Razex's profile picture
Razex3 days ago

where i can see comparaison to this prompt u have some ?

Ishu Agrawal's profile picture
Ishu Agrawal3 days ago

i am so underwhelmed

Bughunter Geek's profile picture
Bughunter Geek3 days ago

Need better checkpoints than that to compete with OpenAI and co.

Juniper's profile picture
Juniper3 days ago

Seems solid

MRLN's profile picture
MRLN3 days ago

on the water ...

Zephyrich's profile picture
Zephyrich3 days ago

When do you estimate it to be released?

Related Videos