Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

🚨 GPT 5.6 Pro first output on the same prompt we are getting started > frontend/ webdev is not solved or improved yet > but understanding increased a lot > it started to take 20-40 mins again like it used to do before 5.5 pro

451,078 Aufrufe • vor 3 Monaten •via X (Twitter)

35 Kommentare

Profilbild von Rijn
Rijnvor 3 Monaten

Damn, it does not look like they solved UI at all. Hopefully this is fake

Profilbild von Zorai
Zoraivor 3 Monaten

Even Glm 5.2 looks way better… really disappointed

Profilbild von Chetaslua
Chetasluavor 3 Monaten

Thanks @mirochill for running prompts and join our server we are cooking there

Profilbild von ρ:ɡeon
ρ:ɡeonvor 3 Monaten

who tf writes all these community notes??

Profilbild von 🥔🥔🥔
🥔🥔🥔vor 3 Monaten

disgusting

Profilbild von Shikhar
Shikharvor 3 Monaten

Frontend isn't solved

Profilbild von Tim B.
Tim B.vor 3 Monaten

Where is 5.6 available

Profilbild von Cory Schulz
Cory Schulzvor 3 Monaten

seems like the anthropic drama had delayed 5.6. i’m not even sure they can release it now. I think it would need to go through that 90 day approval first.

Profilbild von 🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱
🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱vor 3 Monaten

holy shit thats garbage

Profilbild von Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBackvor 3 Monaten

@chetaslua sounds like progress! curious if you've noticed any specific parts of the frontend that are still super annoying or time-consuming?

Profilbild von Kushal
Kushalvor 3 Monaten

maybe they can distill glm 5.2

Profilbild von Yashas
Yashasvor 3 Monaten

Oh this means they focused on engineering

Profilbild von Lex
Lexvor 3 Monaten

I pitted Deepseek v4 flash against the glm 5.2 model for this 3D clock coding task, and strangely enough, Deepseek performed better. I used a simple prompt, and I didn't even write it in English; I wrote it in my own native language.

Profilbild von Choblin
Choblinvor 3 Monaten

svg of ancient clock in 17th century From GPT-5.6 Pro

Profilbild von Adel Bucetta
Adel Bucettavor 3 Monaten

we've seen similar backsliding with other large models too, it's as if they need to relearn the basics every few iterations

Profilbild von Choblin
Choblinvor 3 Monaten

A Pelican Riding a Bicycle Output from GPT-5.6 Pro SVGs are not so Good from this model.

Profilbild von Bryan Patricca
Bryan Patriccavor 3 Monaten

This would be a good one to compare to fable

Profilbild von vibebuilder
vibebuildervor 3 Monaten

It’s always been very strange to me that the models can be incredibly intelligent at back-end, but piss poor on front-end. Like, yes taste is involved, but there’s a lot of logic involved with front-end too.

Profilbild von Nish
Nishvor 3 Monaten

damn gpt 5.6 ugly at UI

Profilbild von brew23
brew23vor 3 Monaten

so not even close to fable

Profilbild von kris.gmg
kris.gmgvor 3 Monaten

looks that fable >>>

Profilbild von Prompt Heat
Prompt Heatvor 3 Monaten

It looks nice. I didn't like the GLM-5.2 design because it was too bright and unrealistic, but this one is cleaner and works more smoothly.

Profilbild von Curline Zephirin
Curline Zephirinvor 3 Monaten

> it started to take 20-40 mins again like it used to do before 5.5 pro @R2Cdev_ Could it be that the one I used that day was GPT 5.6 Pro?

Profilbild von thatguy
thatguyvor 3 Monaten

okay sorry to be one of the people on anthropics side, but 20-40 min for a better result in the low percentages that you can get the same or similar with high fable 5 for a higher price, nothing to use on codex, and still bad ui??? im sorry but fable still my favorite .

Profilbild von Ferbin
Ferbinvor 3 Monaten

sure, understanding improved. but for frontend coding, you traded it for latency. 20-40 min wait and you've forgotten what you were building. was the understanding worth the wait?

Profilbild von omer khan
omer khanvor 3 Monaten

I hope its true

Profilbild von Razex
Razexvor 3 Monaten

Garbage for real bro taking hours for serious project task to be done

Profilbild von J A Z I I
J A Z I Ivor 3 Monaten

Wow this is clean bro

Profilbild von Choblin
Choblinvor 3 Monaten

Why is this case? Why is it now taking same time as previous models? It shouldn't.

Profilbild von InfiniteHexx
InfiniteHexxvor 3 Monaten

If it takes significantly longer, that's proof that 5.6 is a new pre-train

Profilbild von Bughunter Geek
Bughunter Geekvor 3 Monaten

Oh, THAT'S why they said they would solve UI only by the END of the year... 💀

Profilbild von GaleForceAI
GaleForceAIvor 3 Monaten

Yeah, that’s my read too. Better understanding, but frontend still not there. And 20-40 mins is a proper tax.

Profilbild von Tornado guy
Tornado guyvor 3 Monaten

Glm is way better OpenAI should not release this dumb model

Profilbild von Moon 🎑
Moon 🎑vor 3 Monaten

But 5.6 models isn't released yet!?

Profilbild von DrstaOne
DrstaOnevor 3 Monaten

looks same but GPT 5.6 one is better

Ähnliche Videos

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 Aufrufe • vor 3 Monaten