Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Gemini 4 Pro checkpoints have finally started appearing internally a few days ago. This is the first ever output from the model, internally codenamed "argon". It took 2.4 minutes on High thinking effort. It has a 256k token "output limit", compared to 64k in previous Gemini models. I also...

375,337 Aufrufe • vor 4 Tagen •via X (Twitter)

21 Kommentare

Profilbild von Justin
Justinvor 4 Tagen

256k output limit is insane. If they really fix the lazy character of gemini models it might be a cool model!. Output looks okay :)

Profilbild von J A Z I I
J A Z I Ivor 4 Tagen

minimax can do way better then this SIGH

Profilbild von Harshith
Harshithvor 4 Tagen

bro is it really Gemini 4 pro why output looks like this??

Profilbild von Mathis Dupuy
Mathis Dupuyvor 4 Tagen

The bridge is poorly placed samed for the cherry blossom next to the tower … this looks not promising

Profilbild von The Hero of KVcache
The Hero of KVcachevor 4 Tagen

Selling all my goog stock right now

Profilbild von LeMi
LeMivor 4 Tagen

It only took 2.4 minutes, that's a pretty fast model

Profilbild von Louis
Louisvor 4 Tagen

2.4 minutes is quite fast.

Profilbild von LeMi
LeMivor 4 Tagen

is it the RSI model that people keep saying?

Profilbild von Joaoneves
Joaonevesvor 4 Tagen

This is not good, wtf...

Profilbild von Ritwik
Ritwikvor 4 Tagen

256k output is the detail that jumps out. That changes the shape of long-horizon tool traces more than another headline benchmark would.

Profilbild von Calimanu Loredan
Calimanu Loredanvor 4 Tagen

How’s the behavior? Had the chance to test that? Gemini models used to do more, or random stuff even unasked

Profilbild von Ritwik
Ritwikvor 4 Tagen

256K output is the claim that matters if it's a hard generation cap. Context window rumors are cheap until the tokenizer and cache pricing show up.

Profilbild von ethereagle · building
ethereagle · buildingvor 4 Tagen

256k output vs 64k is the number that matters. is that max_tokens, or thinking plus the answer in one budget? High at 2.4 min only helps if reasoning doesn't eat it

Profilbild von lewington
lewingtonvor 4 Tagen

256k output is gonna be really nice for coding agents. you can just give them bigger jobs and let them keep going without having to break everything up so much

Profilbild von VDN
VDNvor 4 Tagen

Amazing news ! Looking forward to it ! I am pretty sure it will be a revolution, because 3.1 pro has aged very well and still is very relevant for many things. Keep up the good work !

Profilbild von Razex
Razexvor 4 Tagen

where i can see comparaison to this prompt u have some ?

Profilbild von Ishu Agrawal
Ishu Agrawalvor 4 Tagen

i am so underwhelmed

Profilbild von Bughunter Geek
Bughunter Geekvor 4 Tagen

Need better checkpoints than that to compete with OpenAI and co.

Profilbild von Juniper
Junipervor 4 Tagen

Seems solid

Profilbild von MRLN
MRLNvor 4 Tagen

on the water ...

Profilbild von Zephyrich
Zephyrichvor 4 Tagen

When do you estimate it to be released?

Ähnliche Videos