Loading video...

Video Failed to Load

Go Home

🚨 GPT 5.6 Pro first output on the same prompt we are getting started > frontend/ webdev is not solved or improved yet > but understanding increased a lot > it started to take 20-40 mins again like it used to do before 5.5 pro

451,078 views • 3 months ago •via X (Twitter)

35 Comments

Rijn's profile picture
Rijn3 months ago

Damn, it does not look like they solved UI at all. Hopefully this is fake

Zorai's profile picture
Zorai3 months ago

Even Glm 5.2 looks way better… really disappointed

Chetaslua's profile picture
Chetaslua3 months ago

Thanks @mirochill for running prompts and join our server we are cooking there

ρ:ɡeon's profile picture
ρ:ɡeon3 months ago

who tf writes all these community notes??

🥔🥔🥔's profile picture
🥔🥔🥔3 months ago

disgusting

Shikhar's profile picture
Shikhar3 months ago

Frontend isn't solved

Tim B.'s profile picture
Tim B.3 months ago

Where is 5.6 available

Cory Schulz's profile picture
Cory Schulz3 months ago

seems like the anthropic drama had delayed 5.6. i’m not even sure they can release it now. I think it would need to go through that 90 day approval first.

🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱's profile picture
🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱3 months ago

holy shit thats garbage

Hussain Hashim | Building SundayBack's profile picture
Hussain Hashim | Building SundayBack3 months ago

@chetaslua sounds like progress! curious if you've noticed any specific parts of the frontend that are still super annoying or time-consuming?

Kushal's profile picture
Kushal3 months ago

maybe they can distill glm 5.2

Yashas's profile picture
Yashas3 months ago

Oh this means they focused on engineering

Lex's profile picture
Lex3 months ago

I pitted Deepseek v4 flash against the glm 5.2 model for this 3D clock coding task, and strangely enough, Deepseek performed better. I used a simple prompt, and I didn't even write it in English; I wrote it in my own native language.

Choblin's profile picture
Choblin3 months ago

svg of ancient clock in 17th century From GPT-5.6 Pro

Adel Bucetta's profile picture
Adel Bucetta3 months ago

we've seen similar backsliding with other large models too, it's as if they need to relearn the basics every few iterations

Choblin's profile picture
Choblin3 months ago

A Pelican Riding a Bicycle Output from GPT-5.6 Pro SVGs are not so Good from this model.

Bryan Patricca's profile picture
Bryan Patricca3 months ago

This would be a good one to compare to fable

vibebuilder's profile picture
vibebuilder3 months ago

It’s always been very strange to me that the models can be incredibly intelligent at back-end, but piss poor on front-end. Like, yes taste is involved, but there’s a lot of logic involved with front-end too.

Nish's profile picture
Nish3 months ago

damn gpt 5.6 ugly at UI

brew23's profile picture
brew233 months ago

so not even close to fable

kris.gmg's profile picture
kris.gmg3 months ago

looks that fable >>>

Prompt Heat's profile picture
Prompt Heat3 months ago

It looks nice. I didn't like the GLM-5.2 design because it was too bright and unrealistic, but this one is cleaner and works more smoothly.

Curline Zephirin's profile picture
Curline Zephirin3 months ago

> it started to take 20-40 mins again like it used to do before 5.5 pro @R2Cdev_ Could it be that the one I used that day was GPT 5.6 Pro?

thatguy's profile picture
thatguy3 months ago

okay sorry to be one of the people on anthropics side, but 20-40 min for a better result in the low percentages that you can get the same or similar with high fable 5 for a higher price, nothing to use on codex, and still bad ui??? im sorry but fable still my favorite .

Ferbin's profile picture
Ferbin3 months ago

sure, understanding improved. but for frontend coding, you traded it for latency. 20-40 min wait and you've forgotten what you were building. was the understanding worth the wait?

omer khan's profile picture
omer khan3 months ago

I hope its true

Razex's profile picture
Razex3 months ago

Garbage for real bro taking hours for serious project task to be done

J A Z I I's profile picture
J A Z I I3 months ago

Wow this is clean bro

Choblin's profile picture
Choblin3 months ago

Why is this case? Why is it now taking same time as previous models? It shouldn't.

InfiniteHexx's profile picture
InfiniteHexx3 months ago

If it takes significantly longer, that's proof that 5.6 is a new pre-train

Bughunter Geek's profile picture
Bughunter Geek3 months ago

Oh, THAT'S why they said they would solve UI only by the END of the year... 💀

GaleForceAI's profile picture
GaleForceAI3 months ago

Yeah, that’s my read too. Better understanding, but frontend still not there. And 20-40 mins is a proper tax.

Tornado guy's profile picture
Tornado guy3 months ago

Glm is way better OpenAI should not release this dumb model

Moon 🎑's profile picture
Moon 🎑3 months ago

But 5.6 models isn't released yet!?

DrstaOne's profile picture
DrstaOne3 months ago

looks same but GPT 5.6 one is better

Related Videos

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 views • 3 months ago