正在加载视频...

视频加载失败

sonnet 5 is routing to sonnet 5.5 it's been a week since it started routing to newer model and team already confirmed that sonnet 5.5 is coming soon from early testing it looks like a huge upgrade from sonnet 5 like opus 5 to opus 5.5 just look at...

17,512 次观看 • 18 小时前 •via X (Twitter)

57 条评论

Deepu 的头像
Deepu18 小时前

is sonnet mogs sol, it's over for openai.

J A Z I I 的头像
J A Z I I17 小时前

They won't let it

Deepanshu Sharma 的头像
Deepanshu Sharma18 小时前

f Sonnet actually mogs Sol in real usage and Ant. drops it around DevDay, then OAI will have to accelerate 6.1 Astra and push an update to the Sol and Luna line like they did with the 5.6 series in august

J A Z I I 的头像
J A Z I I17 小时前

they better drop better sol on dev day

Bee 的头像
Bee18 小时前

When's the 10k post coming!!

J A Z I I 的头像
J A Z I I17 小时前

tomorrow bro

Ferus 的头像
Ferus18 小时前

But what about the use cases? In what context do you use it?

J A Z I I 的头像
J A Z I I18 小时前

just small tests for now gonna have to proper testing soon but yeah

Ferus 的头像
Ferus18 小时前

I abandoned Fable 5.1 completely, it’s super awesome!! I find Opus 5.5 to be much better even in terms of reasoning and logic, and I wonder… what will the next Fable be? A monster?

J A Z I I 的头像
J A Z I I18 小时前

Yeah same for me too next fable would be 5.5 And it will be better then opus 5.5 bit expensive you know

Ferus 的头像
Ferus18 小时前

Yes, but what interests me is a model’s intelligence in terms of investigation, intuition, finding and seeing the non-obvious. That’s where the revolution lies—not just a simple prediction, but changing the game entirely… How exciting the future is… Have you seen Meta’s announcements? I really think Mark is one of the few truly bringing innovation, yet at the same time, in my opinion, very misunderstood! I would really love to have a Muse Charm; the real revolution would be proactivity, with it seeking you out, like a truly living interaction.

Ziwen 的头像
Ziwen17 小时前

No way!! Sonnet 5.5 is goal too haha

J A Z I I 的头像
J A Z I I17 小时前

same reaction bro

Zapupan 的头像
Zapupan17 小时前

6 sol better lol

Eco 的头像
Eco18 小时前

Imagine if sonnet mogs Astra

Vlad Oreshkov 的头像
Vlad Oreshkov17 小时前

If sonnet mogs Sol that’s it. I would be completely switching to Claude. I have 2 codex subscriptions, and Astra is too expensive for plus. If sonnet is this good, I would have no other logical choice.

Mirochill 的头像
Mirochill17 小时前

Imagine Haiku 5.5 > 6 sol 🥲

J A Z I I 的头像
J A Z I I17 小时前

nah that can't be true

Mirochill 的头像
Mirochill17 小时前

I don't hope so lol

Rexei 的头像
Rexei14 小时前

Sonnet seems to be better

Buzz 的头像
Buzz17 小时前

you’re team openai? it’s not looking so good right now lol

J A Z I I 的头像
J A Z I I17 小时前

i am team whoever gives me best model

Tequila Sunset 的头像
Tequila Sunset18 小时前

At least for this, it is not. It's just different style.

Luis Kisters 的头像
Luis Kisters14 小时前

im fkn jealous man. how's sonnet 5? would LOVE to hear more on ur experience working with it. and would be nice if you could do some measurements, re token efficiency and stuff like that

J A Z I I 的头像
J A Z I I12 小时前

ask it do you know tibo the reset guy? Don't search in new chat

Luis Kisters 的头像
Luis Kisters11 小时前

good point shouldve asked directly for codex reset guy or mention tibo. still not sadly but tysm man:)

Intius 的头像
Intius16 小时前

I noticed Anthropic models are always the best in the fields that people publicly test them on. People do tests on games, voxels, 3d, SVGs and that's where it performs the best. I also feel like Astra mogs opus in logic and backend coding which lets me believe Claude 'hypemaxxxes' their models by training them on those specific fields.

Peter Dedene 的头像
Peter Dedene16 小时前

that would be crazy 🤯

J A Z I I 的头像
J A Z I I15 小时前

i am hearing anthropic wanna release this tomorrow

Peter Dedene 的头像
Peter Dedene15 小时前

🔥 they're pacing the other way!

Hybrid 🏃🏾 的头像
Hybrid 🏃🏾18 小时前

Today or tomorrow we may be getting it!!

J A Z I I 的头像
J A Z I I17 小时前

nah maybe in like week or two

Tora Blaze 的头像
Tora Blaze13 小时前

Sol's version makes a great landing page

J A Z I I 的头像
J A Z I I12 小时前

Hmm lemme try

Aimerscoding 的头像
Aimerscoding18 小时前

What....? It looks better than sol .. 😭

J A Z I I 的头像
J A Z I I17 小时前

why it looks better?

Aimerscoding 的头像
Aimerscoding17 小时前

Looks better to me bcz of the level of detail

Zephyr ✪ 的头像
Zephyr ✪18 小时前

Dairo is coming anthropic has been cooking all this while 😂

Alan 的头像
Alan18 小时前

just imagine the anthropic developers

J A Z I I 的头像
J A Z I I17 小时前

I am ant dev

J.𝙳𝚛𝚊𝚟𝚎𝚗 的头像
J.𝙳𝚛𝚊𝚟𝚎𝚗18 小时前

Looking so cool

J A Z I I 的头像
J A Z I I17 小时前

Which one?

J.𝙳𝚛𝚊𝚟𝚎𝚗 的头像
J.𝙳𝚛𝚊𝚟𝚎𝚗17 小时前

Both are but I think sol good

kepo 的头像
kepo17 小时前

Sonnet 5.5 sounds bangoooooor

vengeful181 的头像
vengeful18115 小时前

Imagine being better than opus, as they did it before, lol

J A Z I I 的头像
J A Z I I15 小时前

won't be better then opus but then sonnet 5

vengeful181 的头像
vengeful18115 小时前

Well sonnet models surpassed opus ones before, but this might be too early

Webster | JARVIS 的头像
Webster | JARVIS17 小时前

The routing rumors always make me refresh the playground lol. If 5.5 really closes the gap like that, it would be a nice surprise. Still hoping OpenAI pulls off a comeback too.

Eugene 的头像
Eugene18 小时前

Considering Sol is Terra now and they api pricing is exactly the same, it doesn’t sound that bad, right? I mean Claude models suppose to be better, always been.

Aight Man 的头像
Aight Man17 小时前

tbf 6 'sol' is more like the new terra tbf, I don't think that model deserves same recognition as 5.6 Sol. And the pricing is same as Sonnet 5, so if it beats 6 sol then it won't be too bad since Astra Minor is also coming.

BigJ 的头像
BigJ13 小时前

It's ture, Sonnet now knows Tibo the reset guy without search

J A Z I I 的头像
J A Z I I12 小时前

Only for some people

001q 的头像
001q18 小时前

dm me I need a devin plan... 🫠

Iftikar Alam 的头像
Iftikar Alam17 小时前

Bro gotta check DM once

Edi-san 的头像
Edi-san12 小时前

Like all models bevor, im in the testing pool. I asked sonnet 5.5 the date of his training and he answered end of August 26. Opus 5.5 was june 26 and fable 5.x was July 26 @notjazii

JOY 的头像
JOY16 小时前

Interesting man

max naid 的头像
max naid16 小时前

did they just master post training and are fixing all their models?

相关视频

sonnet 5 vs sonnet 4.6 vs opus 4.8 vs glm 5.2 – frontend tasks dropped sonnet 5 into a quick test today. same three prompts to all four models, single-shot html/canvas, no edits: • objects falling on a trampoline • rockets playing tennis • a slingshot breaking bottles ranked by speed (total across the 3 tasks): 1. opus 4.8 – 15m 09s 2. sonnet 5 – 16m 05s 3. glm 5.2 – 27m 18s 4. sonnet 4.6 – 35m 06s ranked by code shortness (total loc): 1. sonnet 5 – 1794 2. opus 4.8 – 2063 3. sonnet 4.6 – 2182 4. glm 5.2 – 3285 sonnet 5 came out on top here – leanest code overall and a near-tie for fastest it was also the most creative. in every task it added something none of the others did: – kept the trampoline vibrating after the objects landed – drew a +1 next to the rocket that scored the point – turned the slingshot to face the next bottle before each shot opus 4.8 evaluated the code sonnet 5 produced. four things stood out: • the sphere is a fake, and that's the smart move. the cube and star are real 3d meshes with proper culling and shading, but the ball is just a flat shaded circle. a lit sphere looks identical from every angle, so building it in 3d would burn compute for zero visible payoff. knowing where not to bother is its own kind of skill • weight actually means something on the trampoline. the star is heavy, so it barely bounces and dents the mat hard. the ball is light, so it's lively and leaves a shallow dip. the three objects aren't just different shapes – they have different temperaments, and the physics is what gives them that • the slingshot is framed like a shot, not just drawn. the handle is anchored below the bottom of the screen and runs off-frame, so it reads as something you're holding rather than a sprite parked in the scene. that's a staging instinct, not a rendering one • the paddle ai forward-simulates the ball to predict where it'll land, then adds a deliberate error bias (roughly 1 in 5 shots is a real miss). that's why scoring looks natural instead of robotic – plus four distinct fault types with a catch-all so a rally never hangs without a result bottom line: sonnet 5 does more with less. fastest tier, leanest code, and the only one that added small touches nobody asked for follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

14,518 次观看 • 2 个月前