Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

🚨 GPT 5.6 Pro first output on the same prompt we are getting started > frontend/ webdev is not solved or improved yet > but understanding increased a lot > it started to take 20-40 mins again like it used to do before 5.5 pro

451,078 görüntüleme • 3 ay önce •via X (Twitter)

35 Yorum

Rijn profil fotoğrafı
Rijn3 ay önce

Damn, it does not look like they solved UI at all. Hopefully this is fake

Zorai profil fotoğrafı
Zorai3 ay önce

Even Glm 5.2 looks way better… really disappointed

Chetaslua profil fotoğrafı
Chetaslua3 ay önce

Thanks @mirochill for running prompts and join our server we are cooking there

ρ:ɡeon profil fotoğrafı
ρ:ɡeon3 ay önce

who tf writes all these community notes??

🥔🥔🥔 profil fotoğrafı
🥔🥔🥔3 ay önce

disgusting

Shikhar profil fotoğrafı
Shikhar3 ay önce

Frontend isn't solved

Tim B. profil fotoğrafı
Tim B.3 ay önce

Where is 5.6 available

Cory Schulz profil fotoğrafı
Cory Schulz3 ay önce

seems like the anthropic drama had delayed 5.6. i’m not even sure they can release it now. I think it would need to go through that 90 day approval first.

🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱 profil fotoğrafı
🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱3 ay önce

holy shit thats garbage

Hussain Hashim | Building SundayBack profil fotoğrafı
Hussain Hashim | Building SundayBack3 ay önce

@chetaslua sounds like progress! curious if you've noticed any specific parts of the frontend that are still super annoying or time-consuming?

Kushal profil fotoğrafı
Kushal3 ay önce

maybe they can distill glm 5.2

Yashas profil fotoğrafı
Yashas3 ay önce

Oh this means they focused on engineering

Lex profil fotoğrafı
Lex3 ay önce

I pitted Deepseek v4 flash against the glm 5.2 model for this 3D clock coding task, and strangely enough, Deepseek performed better. I used a simple prompt, and I didn't even write it in English; I wrote it in my own native language.

Choblin profil fotoğrafı
Choblin3 ay önce

svg of ancient clock in 17th century From GPT-5.6 Pro

Adel Bucetta profil fotoğrafı
Adel Bucetta3 ay önce

we've seen similar backsliding with other large models too, it's as if they need to relearn the basics every few iterations

Choblin profil fotoğrafı
Choblin3 ay önce

A Pelican Riding a Bicycle Output from GPT-5.6 Pro SVGs are not so Good from this model.

Bryan Patricca profil fotoğrafı
Bryan Patricca3 ay önce

This would be a good one to compare to fable

vibebuilder profil fotoğrafı
vibebuilder3 ay önce

It’s always been very strange to me that the models can be incredibly intelligent at back-end, but piss poor on front-end. Like, yes taste is involved, but there’s a lot of logic involved with front-end too.

Nish profil fotoğrafı
Nish3 ay önce

damn gpt 5.6 ugly at UI

brew23 profil fotoğrafı
brew233 ay önce

so not even close to fable

kris.gmg profil fotoğrafı
kris.gmg3 ay önce

looks that fable >>>

Prompt Heat profil fotoğrafı
Prompt Heat3 ay önce

It looks nice. I didn't like the GLM-5.2 design because it was too bright and unrealistic, but this one is cleaner and works more smoothly.

Curline Zephirin profil fotoğrafı
Curline Zephirin3 ay önce

> it started to take 20-40 mins again like it used to do before 5.5 pro @R2Cdev_ Could it be that the one I used that day was GPT 5.6 Pro?

thatguy profil fotoğrafı
thatguy3 ay önce

okay sorry to be one of the people on anthropics side, but 20-40 min for a better result in the low percentages that you can get the same or similar with high fable 5 for a higher price, nothing to use on codex, and still bad ui??? im sorry but fable still my favorite .

Ferbin profil fotoğrafı
Ferbin3 ay önce

sure, understanding improved. but for frontend coding, you traded it for latency. 20-40 min wait and you've forgotten what you were building. was the understanding worth the wait?

omer khan profil fotoğrafı
omer khan3 ay önce

I hope its true

Razex profil fotoğrafı
Razex3 ay önce

Garbage for real bro taking hours for serious project task to be done

J A Z I I profil fotoğrafı
J A Z I I3 ay önce

Wow this is clean bro

Choblin profil fotoğrafı
Choblin3 ay önce

Why is this case? Why is it now taking same time as previous models? It shouldn't.

InfiniteHexx profil fotoğrafı
InfiniteHexx3 ay önce

If it takes significantly longer, that's proof that 5.6 is a new pre-train

Bughunter Geek profil fotoğrafı
Bughunter Geek3 ay önce

Oh, THAT'S why they said they would solve UI only by the END of the year... 💀

GaleForceAI profil fotoğrafı
GaleForceAI3 ay önce

Yeah, that’s my read too. Better understanding, but frontend still not there. And 20-40 mins is a proper tax.

Tornado guy profil fotoğrafı
Tornado guy3 ay önce

Glm is way better OpenAI should not release this dumb model

Moon 🎑 profil fotoğrafı
Moon 🎑3 ay önce

But 5.6 models isn't released yet!?

DrstaOne profil fotoğrafı
DrstaOne3 ay önce

looks same but GPT 5.6 one is better

Benzer Videolar

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 görüntüleme • 3 ay önce