Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Which flash model wins? DeepSeek v4.1 flash, Qwen 3.8 flash, or GLM 5.3 flash 🤯

15,052 görüntüleme • 4 gün önce •via X (Twitter)

13 Yorum

Praveen Yadav profil fotoğrafı
Praveen Yadav4 gün önce

Glm 5.3 flash looks like a winner to me

VulKan profil fotoğrafı
VulKan4 gün önce

4.1 Flash takes the cake

Uday G profil fotoğrafı
Uday G4 gün önce

i liked the deepseek v4.1 result

Konstantin Anagnostou profil fotoğrafı
Konstantin Anagnostou4 gün önce

DS for sure

Maحmoud Ismaعil profil fotoğrafı
Maحmoud Ismaعil4 gün önce

Qwen

SK profil fotoğrafı
SK4 gün önce

Why did y need to add V4 pro ? U comparing flash models

Alina Ai profil fotoğrafı
Alina Ai4 gün önce

DeepSeek V4.1 Flash looks like the one to beat right now.

Khen profil fotoğrafı
Khen4 gün önce

Very cool test

GMI Cloud profil fotoğrafı
GMI Cloud4 gün önce

thanks!

Kawai 🎀 profil fotoğrafı
Kawai 🎀4 gün önce

4.1 flashhh it's cool

sigma profil fotoğrafı
sigma4 gün önce

Which harness was this on?

Ahmet profil fotoğrafı
Ahmet4 gün önce

i see 4 claude models

Rupert profil fotoğrafı
Rupert4 gün önce

Please give this idea a try.... 1. Take the code from all of them, merge into zip or document. 2. Give all of them the merged file and do a second pass to see how they adapt to the new information 3. Show us ❤️

Benzer Videolar

glm 5.3 vs qwen 3.8 vs gemini 3.7 vs deepseek v4 flash four models designed and built three structures each on a physics-backed site, with no dimensions anywhere in the brief the setup: our own agent loop on OpenRouter, a construction site as the tool set – footings, walls, arches, roofs, scaffold, a lamp. the site enforces physics and nothing else: unsupported brick falls, a roof needs walls under it, a worker reaches 3.2 m above whatever he stands on, an arch needs centring until the keystone is set, concrete cures before it carries. no budget ceiling – material cost is tallied and reported, never blocked. tasks: 1. house – a plot and a palette, no plan. shape, height and material are the model's call 2. lighthouse – a headland cut by a gully, with a rock stack standing 30 m offshore. the lamp must burn, it must be the highest thing built, and the keeper must be able to walk to it 3. bridge – a river with one islet and banks at different heights. cross it however you want models: Z.ai glm 5.3 flash, Qwen qwen 3.8 flash, Google DeepMind gemini 3.7 flash, DeepSeek v4 flash vision all twelve objects were finished and signed off by the models themselves. tallest lighthouse is qwen's at 38.4 m, planted on the offshore stack with a bridge run out to it – the only model that read the site that way. deepseek signed off its bridge on an empty riverbed: 0 bricks, 107 minutes, $1.16m of material tallied - total cost, three builds #1 glm 5.3 flash – $0.201 #2 gemini 3.7 flash – $0.871 #3 qwen 3.8 flash – $1.058 #4 deepseek v4 flash – $1.567 - wall clock, three builds #1 gemini 3.7 flash – 91m #2 glm 5.3 flash – 228m #3 deepseek v4 flash – 502m #4 qwen 3.8 flash – 912m - total tokens #1 gemini 3.7 flash – 3,567,052 #2 glm 5.3 flash – 4,732,748 #3 qwen 3.8 flash – 13,469,333 #4 deepseek v4 flash – 18,230,076 - defects logged by the site #1 deepseek v4 flash – 59 #2 gemini 3.7 flash – 132 #3 glm 5.3 flash – 221 #4 qwen 3.8 flash – 350 - material tallied across three builds #1 gemini 3.7 flash – $359,884 #2 glm 5.3 flash – $583,358 #3 deepseek v4 flash – $1,327,484 #4 qwen 3.8 flash – $2,188,625 observations: • glm is the cheap one and nothing here is close – $0.201 for three buildings, $0.042 per million tokens, 6x under gemini's rate • what glm spends it on is bulk, not care: 166,228 bricks in one house and 156 defect weight, the worst single object in the set • gemini is the efficiency line – 91 minutes and 3.57m tokens for all three and an eighth of qwen's clock • gemini also builds the smallest of everything. its lighthouse is 22.5 m against qwen's 38.4, its house 6.9 m against 19.3 • qwen is the maximalist: 1.18m bricks, $2.19m of material, tallest on all three tasks, and 912 minutes – 15 hours – to get there conclusion: twelve finished objects for $3.80 all in, and a 7.8x price spread between the cheapest model and the priciest! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

26,250 görüntüleme • 18 gün önce