Загрузка видео...

Не удалось загрузить видео

На главную

Claude Opus 5.5 completely MOGS GPT 6 Sol. Same prompt. Same task. Look at the results. Opus 5.5 has way better taste and design. It is not subtle. You can see it immediately. GPT 6 Sol is a step up from GPT 5.6 Sol. But next to Opus 5.5...

55,238 просмотров • 7 дней назад •via X (Twitter)

Комментарии: 42

Фото профиля DeakSpeaks
DeakSpeaks7 дней назад

GPT 6 Sol is half the price of Opus 5.5. You never mentioned that, why not? One is literally half the cost of the other, whilst offering comparable output.

Фото профиля Bridgebench
Bridgebench7 дней назад

Fair point, cost matters too

Фото профиля Rubens Soto | AI & SaaS
Rubens Soto | AI & SaaS7 дней назад

Man, since I'm not creating a FPS game and I need things done with a good ratio of price and performence I prefer gpt 6 sol.

Фото профиля Kaleb Campbell
Kaleb Campbell7 дней назад

Devastating to say the least…

Фото профиля Absurd
Absurd7 дней назад

Sol looks better imo, the only difference is Sol had a pistol and Opus had an automatic. Sol is also half the price, so this doesn’t seem like a great comparison, but it looked like Sol won to me.

Фото профиля pro gamer
pro gamer7 дней назад

Opus costs 5x more and took 6x longer, so this comparison doesn’t really make sense. Opus 5.5 costs roughly 2× more per task than Astra, so that would be a much more reasonable comparison.

Фото профиля Bridgebench
Bridgebench7 дней назад

Fair, speed and cost count too

Фото профиля Dr. Dennis Griffin
Dr. Dennis Griffin7 дней назад

Curious as to why you use GPT 6 Sol and not GPT 6 Astra? And what effort level for each (max or xtra high?), the total cost, total tokens and total time it took to create. Then I can make a better determination. Details please. Thanks.

Фото профиля Mikelchocano 🇪🇸 🛠️
Mikelchocano 🇪🇸 🛠️7 дней назад

What about Opus 5.5 vs Gpt 6 Astra?

Фото профиля Forrest MacDougall
Forrest MacDougall7 дней назад

I think if you take into account speed/cost, they both have a place in your work flow. I do plenty of boring work, and Sol will satisfy that end much better than Opus.

Фото профиля Bridgebench
Bridgebench7 дней назад

Agreed, both have their lane

Фото профиля Saksham Saini
Saksham Saini7 дней назад

Opus cooked as always

Фото профиля Bridgebench
Bridgebench7 дней назад

It really did on this one

Фото профиля Tushar
Tushar7 дней назад

Anthropic still holds the one-shot crown. 👑

Фото профиля Bridgebench
Bridgebench7 дней назад

For now at least

Фото профиля Aditya Shelke
Aditya Shelke7 дней назад

The way Anthropic plays with OpenAI is funny💀

Фото профиля Anthony Aguilar
Anthony Aguilar7 дней назад

All we care about is improvement and cost. We got both today, that’s a W. We are at the lowest point AI will ever be. Think about that.

Фото профиля Bridgebench
Bridgebench7 дней назад

Big W for everyone today

Фото профиля NoneLet
NoneLet7 дней назад

yeah, the shit yellow crab absolutely smoked GPT here. frontend taste carries through everything, the overall style, the detail in the models, the textures, even that “soul” artists keep going on about.

Фото профиля Webster | JARVIS
Webster | JARVIS7 дней назад

This side-by-side is so helpful. Taste is hard to quantify, but you can feel it right away. Curious if it stays good on messy real-world prompts too.

Фото профиля Bridgebench
Bridgebench7 дней назад

Hard to benchmark, easy to see

Фото профиля bittu
bittu7 дней назад

Bro Opus 5.5 is the best. I hope in subscription it's good usage as well and doesn't feel like fable.

Фото профиля Bridgebench
Bridgebench7 дней назад

Same, usage limits will decide it

Фото профиля bittu
bittu7 дней назад

True

Фото профиля Iván Lanchazo
Iván Lanchazo7 дней назад

El mismo prompt no mide el gusto. Mide qué modelo acierta más con el gusto de quien escribió ese prompt. La comparación útil empieza cuando defines criterios antes de ver el resultado.

Фото профиля Vadim Pavlov
Vadim Pavlov7 дней назад

I’m rooting for OpenAI, but let’s be honest—they fell flat on their face today.

Фото профиля J A Z I I
J A Z I I7 дней назад

Test same level model lol 😆

Фото профиля Eco
Eco7 дней назад

Sol is not even close to opus

Фото профиля YogenshaSilver
YogenshaSilver7 дней назад

Slop Bench at it again

Фото профиля The Bullish Broke Guy X
The Bullish Broke Guy X7 дней назад

Same prompt, same task, same benchmark vendor incentives. The real test is not which model wins the demo, it is which one is still the cheapest per correct answer six months from now.

Фото профиля Bak
Bak7 дней назад

Try this on Astra Ultra Max PLEASE

Фото профиля LeMi
LeMi7 дней назад

On average, GPT-6 Sol is 5.6x cheaper than Opus 5.5 for the task based on AA

Фото профиля Mauri
Mauri7 дней назад

unblock @argofowl, he comes in peace

Фото профиля BlanPlan
BlanPlan7 дней назад

Same prompt is a good control. I'd want to see it hold across a few different briefs before calling it taste — one task is where that usually falls apart.

Фото профиля LoiEtang
LoiEtang7 дней назад

the show case yu do are built in web game/ simulation which will never be used for big project, gpt 6 for exemple is so much better at using UE5 right now than opus 5.5

Фото профиля Beto.tsx
Beto.tsx7 дней назад

mind sharing prompt and resources you used to create it?

Фото профиля openaiteach
openaiteach7 дней назад

What is the prompt

Фото профиля Marcelo Retana
Marcelo Retana7 дней назад

Opus 5.5 is only "good" at delivering UI stuff. It's like a mid-senior level engineer. The real deal is GPT Sol/Astra. You can get a lot more done with it.

Фото профиля Toni
Toni7 дней назад

@bridgemindai Does it still feel lazy and slop?

Фото профиля Ben cooper 🇪🇺
Ben cooper 🇪🇺7 дней назад

Not the same price idiot 🤦‍♂️

Фото профиля 刘朝 Zhao Liu
刘朝 Zhao Liu7 дней назад

A visual comparison is useful when the rubric is fixed before the outputs are seen. Measure hierarchy, spacing, typography, responsive behavior, and edit distance across the same prompt and assets; “better taste” becomes a reproducible result instead of a creator preference.

Фото профиля Kotlin
Kotlin7 дней назад

Was the same amount of money used for both ?

Похожие видео