正在加载视频...

视频加载失败

Gemini 4 Pro checkpoints have finally started appearing internally a few days ago. This is the first ever output from the model, internally codenamed "argon". It took 2.4 minutes on High thinking effort. It has a 256k token "output limit", compared to 64k in previous Gemini models. I also...

375,337 次观看 • 4 天前 •via X (Twitter)

21 条评论

Justin 的头像
Justin4 天前

256k output limit is insane. If they really fix the lazy character of gemini models it might be a cool model!. Output looks okay :)

J A Z I I 的头像
J A Z I I4 天前

minimax can do way better then this SIGH

Harshith 的头像
Harshith4 天前

bro is it really Gemini 4 pro why output looks like this??

Mathis Dupuy 的头像
Mathis Dupuy4 天前

The bridge is poorly placed samed for the cherry blossom next to the tower … this looks not promising

The Hero of KVcache 的头像
The Hero of KVcache4 天前

Selling all my goog stock right now

LeMi 的头像
LeMi4 天前

It only took 2.4 minutes, that's a pretty fast model

Louis 的头像
Louis4 天前

2.4 minutes is quite fast.

LeMi 的头像
LeMi4 天前

is it the RSI model that people keep saying?

Joaoneves 的头像
Joaoneves4 天前

This is not good, wtf...

Ritwik 的头像
Ritwik4 天前

256k output is the detail that jumps out. That changes the shape of long-horizon tool traces more than another headline benchmark would.

Calimanu Loredan 的头像
Calimanu Loredan4 天前

How’s the behavior? Had the chance to test that? Gemini models used to do more, or random stuff even unasked

Ritwik 的头像
Ritwik4 天前

256K output is the claim that matters if it's a hard generation cap. Context window rumors are cheap until the tokenizer and cache pricing show up.

ethereagle · building 的头像
ethereagle · building4 天前

256k output vs 64k is the number that matters. is that max_tokens, or thinking plus the answer in one budget? High at 2.4 min only helps if reasoning doesn't eat it

lewington 的头像
lewington4 天前

256k output is gonna be really nice for coding agents. you can just give them bigger jobs and let them keep going without having to break everything up so much

VDN 的头像
VDN4 天前

Amazing news ! Looking forward to it ! I am pretty sure it will be a revolution, because 3.1 pro has aged very well and still is very relevant for many things. Keep up the good work !

Razex 的头像
Razex4 天前

where i can see comparaison to this prompt u have some ?

Ishu Agrawal 的头像
Ishu Agrawal4 天前

i am so underwhelmed

Bughunter Geek 的头像
Bughunter Geek4 天前

Need better checkpoints than that to compete with OpenAI and co.

Juniper 的头像
Juniper4 天前

Seems solid

MRLN 的头像
MRLN4 天前

on the water ...

Zephyrich 的头像
Zephyrich4 天前

When do you estimate it to be released?

相关视频