Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

🚨 Big Week Ahead on one side Millennium Prize Problems ( they are contacting mathematicians will give few citations to them in solution to make them happy this time) for consumer side > fable 5.2 > opus 5.2 , Sol 6 > grok 4.7 and gemini 4 pro if...

26,767 görüntüleme • 1 gün önce •via X (Twitter)

16 Yorum

lyra profil fotoğrafı
lyra1 gün önce

@notjazii

Afterkind profil fotoğrafı
Afterkind1 gün önce

Accelerating looks like ASI in 3 months. OpenAI could theoretically say "fuck everything, we know this will work", shut down chatgpt and codex and dedicate their entire compute for a massive training run and having Bel optimize the entire thing until a 50T behemoth comes out without safety guardrails. That would be the "fuck it, we're accelerating" case.

刘朝 Zhao Liu profil fotoğrafı
刘朝 Zhao Liu1 gün önce

The release pile-up is exciting, yet routing and access labels make the comparison slippery. A stable model ID, cohort trace, and fixed replay set would show whether Gemini 4 Pro is a new checkpoint or a serving-path surprise.

lost in latency profil fotoğrafı
lost in latency1 gün önce

is 6 sol oai's answer to fable 5.2 and opus 5.2?

Chetaslua profil fotoğrafı
Chetaslua1 gün önce

no it is the question actually

lost in latency profil fotoğrafı
lost in latency1 gün önce

yeah my question as well

刘朝 Zhao Liu profil fotoğrafı
刘朝 Zhao Liu1 gün önce

The release cadence is accelerating, but the useful comparison needs a stable harness. Track model identity, routing label, tool budget, latency, and recovery on the same long task; otherwise a packed launch week produces headlines without a durable frontier map.

Marsonal profil fotoğrafı
Marsonal1 gün önce

Gemini 4 releases next week?

Shesaidmewakeup profil fotoğrafı
Shesaidmewakeup1 gün önce

If all of those land in one week, pacing is just the press word for a traffic jam. The consumer side will feel it first.

AlLeakWire profil fotoğrafı
AlLeakWire1 gün önce

Bro is gemini 4 pro comparable to the opus 5.2 and gpt 6 astra

Chetaslua profil fotoğrafı
Chetaslua1 gün önce

i will get hate but gemini 4 pro is bad at tool calling you are listening this from me first , but in my experience in last two days this model forget he is in code sandbox env

AlLeakWire profil fotoğrafı
AlLeakWire1 gün önce

Is that the model confirms to be gemini 4

Chetaslua profil fotoğrafı
Chetaslua1 gün önce

this is what our majority consensus in discord community

AlLeakWire profil fotoğrafı
AlLeakWire1 gün önce

Means it was not the google comeback this year

Chetaslua profil fotoğrafı
Chetaslua1 gün önce

its good google model , best in frontend and one shot demo and day to day use

Samuel Hu profil fotoğrafı
Samuel Hu1 gün önce

the routed model path is the useful comparison. i’d keep one ugly repo task fixed across all three, then compare diff quality and recovery, not just the max-setting score

Benzer Videolar