Загрузка видео...

Не удалось загрузить видео

На главную

🚨 GPT 5.6 Pro first output on the same prompt we are getting started > frontend/ webdev is not solved or improved yet > but understanding increased a lot > it started to take 20-40 mins again like it used to do before 5.5 pro

451,078 просмотров • 3 месяцев назад •via X (Twitter)

Комментарии: 35

Фото профиля Rijn
Rijn3 месяцев назад

Damn, it does not look like they solved UI at all. Hopefully this is fake

Фото профиля Zorai
Zorai3 месяцев назад

Even Glm 5.2 looks way better… really disappointed

Фото профиля Chetaslua
Chetaslua3 месяцев назад

Thanks @mirochill for running prompts and join our server we are cooking there

Фото профиля ρ:ɡeon
ρ:ɡeon3 месяцев назад

who tf writes all these community notes??

Фото профиля 🥔🥔🥔
🥔🥔🥔3 месяцев назад

disgusting

Фото профиля Shikhar
Shikhar3 месяцев назад

Frontend isn't solved

Фото профиля Tim B.
Tim B.3 месяцев назад

Where is 5.6 available

Фото профиля Cory Schulz
Cory Schulz3 месяцев назад

seems like the anthropic drama had delayed 5.6. i’m not even sure they can release it now. I think it would need to go through that 90 day approval first.

Фото профиля 🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱
🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱🇵🇱3 месяцев назад

holy shit thats garbage

Фото профиля Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBack3 месяцев назад

@chetaslua sounds like progress! curious if you've noticed any specific parts of the frontend that are still super annoying or time-consuming?

Фото профиля Kushal
Kushal3 месяцев назад

maybe they can distill glm 5.2

Фото профиля Yashas
Yashas3 месяцев назад

Oh this means they focused on engineering

Фото профиля Lex
Lex3 месяцев назад

I pitted Deepseek v4 flash against the glm 5.2 model for this 3D clock coding task, and strangely enough, Deepseek performed better. I used a simple prompt, and I didn't even write it in English; I wrote it in my own native language.

Фото профиля Choblin
Choblin3 месяцев назад

svg of ancient clock in 17th century From GPT-5.6 Pro

Фото профиля Adel Bucetta
Adel Bucetta3 месяцев назад

we've seen similar backsliding with other large models too, it's as if they need to relearn the basics every few iterations

Фото профиля Choblin
Choblin3 месяцев назад

A Pelican Riding a Bicycle Output from GPT-5.6 Pro SVGs are not so Good from this model.

Фото профиля Bryan Patricca
Bryan Patricca3 месяцев назад

This would be a good one to compare to fable

Фото профиля vibebuilder
vibebuilder3 месяцев назад

It’s always been very strange to me that the models can be incredibly intelligent at back-end, but piss poor on front-end. Like, yes taste is involved, but there’s a lot of logic involved with front-end too.

Фото профиля Nish
Nish3 месяцев назад

damn gpt 5.6 ugly at UI

Фото профиля brew23
brew233 месяцев назад

so not even close to fable

Фото профиля kris.gmg
kris.gmg3 месяцев назад

looks that fable >>>

Фото профиля Prompt Heat
Prompt Heat3 месяцев назад

It looks nice. I didn't like the GLM-5.2 design because it was too bright and unrealistic, but this one is cleaner and works more smoothly.

Фото профиля Curline Zephirin
Curline Zephirin3 месяцев назад

> it started to take 20-40 mins again like it used to do before 5.5 pro @R2Cdev_ Could it be that the one I used that day was GPT 5.6 Pro?

Фото профиля thatguy
thatguy3 месяцев назад

okay sorry to be one of the people on anthropics side, but 20-40 min for a better result in the low percentages that you can get the same or similar with high fable 5 for a higher price, nothing to use on codex, and still bad ui??? im sorry but fable still my favorite .

Фото профиля Ferbin
Ferbin3 месяцев назад

sure, understanding improved. but for frontend coding, you traded it for latency. 20-40 min wait and you've forgotten what you were building. was the understanding worth the wait?

Фото профиля omer khan
omer khan3 месяцев назад

I hope its true

Фото профиля Razex
Razex3 месяцев назад

Garbage for real bro taking hours for serious project task to be done

Фото профиля J A Z I I
J A Z I I3 месяцев назад

Wow this is clean bro

Фото профиля Choblin
Choblin3 месяцев назад

Why is this case? Why is it now taking same time as previous models? It shouldn't.

Фото профиля InfiniteHexx
InfiniteHexx3 месяцев назад

If it takes significantly longer, that's proof that 5.6 is a new pre-train

Фото профиля Bughunter Geek
Bughunter Geek3 месяцев назад

Oh, THAT'S why they said they would solve UI only by the END of the year... 💀

Фото профиля GaleForceAI
GaleForceAI3 месяцев назад

Yeah, that’s my read too. Better understanding, but frontend still not there. And 20-40 mins is a proper tax.

Фото профиля Tornado guy
Tornado guy3 месяцев назад

Glm is way better OpenAI should not release this dumb model

Фото профиля Moon 🎑
Moon 🎑3 месяцев назад

But 5.6 models isn't released yet!?

Фото профиля DrstaOne
DrstaOne3 месяцев назад

looks same but GPT 5.6 one is better

Похожие видео

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 просмотров • 3 месяцев назад