Loading video...

Video Failed to Load

Go Home

new Gemini 4 Pro checkpoint is here and is looking a lot better this car prompt took roughly 2-3 mins, and is being ghost tested under 3.8 flash it’s a random encounter when in battle mode on arena ai using 3.8 flash and you can only judge by the...

49,021 views • 3 days ago •via X (Twitter)

18 Comments

IDK's profile picture
IDK3 days ago

It's a good thing, but... Let's compare it with real Gemini 3.8 flash vs Gemini 4 (this checkpoint) How clear are the differences?

Qwinah's profile picture
Qwinah3 days ago

3.8 flash couldn’t create this quality of render, you are right tho I will get a 3.8 flash version to showcase this

Singularity's profile picture
Singularity3 days ago

Gemini 4 Pro

Qwinah's profile picture
Qwinah3 days ago

You reckon? Just found its internal name to be Argon

Singularity's profile picture
Singularity3 days ago

Maybe, but it still feels like close to Fable 5 level.

J A Z I I's profile picture
J A Z I I3 days ago

bro is cooking harder

Qwinah's profile picture
Qwinah3 days ago

harder than ever bro 😂

AKHIL's profile picture
AKHIL3 days ago

I don't think so it's gemini 4

Qwinah's profile picture
Qwinah3 days ago

Just found the internal code name being Argon

CruxLog's profile picture
CruxLog3 days ago

If this is gemini 4 then it's again another disappointment, astra and fable 5.1 can generate better output then this, so again it's not gonna be frontier.

Ke's profile picture
Ke3 days ago

Lego? 🤡

tolgaozisik's profile picture
tolgaozisik3 days ago

this is clearly 3.9 flash

Qwinah's profile picture
Qwinah3 days ago

What makes you say that?

tolgaozisik's profile picture
tolgaozisik3 days ago

I know some model behaviours and this is strongly flash model. Gemini 4.0 will be really different ground up.

Sophia's profile picture
Sophia3 days ago

Seems like 3.9 flash not 4 pro

xalid's profile picture
xalid3 days ago

Its not Gemini 4 pro

Qwinah's profile picture
Qwinah3 days ago

How not?

Javier Jiménez's profile picture
Javier Jiménez3 days ago

Uy si Fanboy de Google espera que salga al público y será una mierda como todo su modelo anteriores

Related Videos

glm 5.3 vs qwen 3.8 vs gemini 3.7 vs deepseek v4 flash four models designed and built three structures each on a physics-backed site, with no dimensions anywhere in the brief the setup: our own agent loop on OpenRouter, a construction site as the tool set – footings, walls, arches, roofs, scaffold, a lamp. the site enforces physics and nothing else: unsupported brick falls, a roof needs walls under it, a worker reaches 3.2 m above whatever he stands on, an arch needs centring until the keystone is set, concrete cures before it carries. no budget ceiling – material cost is tallied and reported, never blocked. tasks: 1. house – a plot and a palette, no plan. shape, height and material are the model's call 2. lighthouse – a headland cut by a gully, with a rock stack standing 30 m offshore. the lamp must burn, it must be the highest thing built, and the keeper must be able to walk to it 3. bridge – a river with one islet and banks at different heights. cross it however you want models: Z.ai glm 5.3 flash, Qwen qwen 3.8 flash, Google DeepMind gemini 3.7 flash, DeepSeek v4 flash vision all twelve objects were finished and signed off by the models themselves. tallest lighthouse is qwen's at 38.4 m, planted on the offshore stack with a bridge run out to it – the only model that read the site that way. deepseek signed off its bridge on an empty riverbed: 0 bricks, 107 minutes, $1.16m of material tallied - total cost, three builds #1 glm 5.3 flash – $0.201 #2 gemini 3.7 flash – $0.871 #3 qwen 3.8 flash – $1.058 #4 deepseek v4 flash – $1.567 - wall clock, three builds #1 gemini 3.7 flash – 91m #2 glm 5.3 flash – 228m #3 deepseek v4 flash – 502m #4 qwen 3.8 flash – 912m - total tokens #1 gemini 3.7 flash – 3,567,052 #2 glm 5.3 flash – 4,732,748 #3 qwen 3.8 flash – 13,469,333 #4 deepseek v4 flash – 18,230,076 - defects logged by the site #1 deepseek v4 flash – 59 #2 gemini 3.7 flash – 132 #3 glm 5.3 flash – 221 #4 qwen 3.8 flash – 350 - material tallied across three builds #1 gemini 3.7 flash – $359,884 #2 glm 5.3 flash – $583,358 #3 deepseek v4 flash – $1,327,484 #4 qwen 3.8 flash – $2,188,625 observations: • glm is the cheap one and nothing here is close – $0.201 for three buildings, $0.042 per million tokens, 6x under gemini's rate • what glm spends it on is bulk, not care: 166,228 bricks in one house and 156 defect weight, the worst single object in the set • gemini is the efficiency line – 91 minutes and 3.57m tokens for all three and an eighth of qwen's clock • gemini also builds the smallest of everything. its lighthouse is 22.5 m against qwen's 38.4, its house 6.9 m against 19.3 • qwen is the maximalist: 1.18m bricks, $2.19m of material, tallest on all three tasks, and 912 minutes – 15 hours – to get there conclusion: twelve finished objects for $3.80 all in, and a 7.8x price spread between the cheapest model and the priciest! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

26,360 views • 22 days ago