正在加载视频...

视频加载失败

We’ve upgraded our specialized reasoning mode Gemini 3 Deep Think to help solve modern science, research, and engineering challenges – pushing the frontier of intelligence. 🧠 Watch how the Wang Lab at Duke University is using it to design new semiconductor materials. 🧵

3,200,343 次观看 • 7 个月前 •via X (Twitter)

35 条评论

Google DeepMind 的头像
Google DeepMind7 个月前

The latest Deep Think moves beyond abstract theory to drive practical applications. It’s state-of-the-art on ARC-AGI-2, a benchmark for frontier AI reasoning. On Humanity’s Last Exam, it sets a new standard, tackling the hardest problems across mathematics, science, and engineering — making it a genuine collaborator for heavy-duty analysis. It achieved an Elo of 3455 on Codeforces, demonstrating the ability to solve complex, real-world coding tasks - while earning gold medal-level results on the written portion of the 2025 Physics and Chemistry Olympiads.

Google DeepMind 的头像
Google DeepMind7 个月前

The upgraded Deep Think mode is rolling out now in the @GeminiApp for Google AI Ultra subscribers. For scientific researchers and developers, we’re opening a Vertex AI Early Access Program for the API. Start discovering →

Marko Njegomir 的头像
Marko Njegomir7 个月前

84.6% on ARC-AGI 2, and it's only February of 2026. Google cooked.

Besci 的头像
Besci7 个月前

Model companies waiting for their competitor to release a new model so they can release theirs a day later and steal the news cycle.

Leon Lin 的头像
Leon Lin7 个月前

damn that is not a little upgrade guys

A.Bernardi 的头像
A.Bernardi7 个月前

brutal frame mog for gptcels holy cortisol spike for opuscels giga lifefuel for geminicels over for arc-agi 2 benchmarkcels never began for "the wall" copers

Dharmik 的头像
Dharmik7 个月前

Damn, they mogged so hard!!

Patel Meet 的头像
Patel Meet7 个月前

crazy benchmarks!!!!!

Gaffar Al-Ansari 的头像
Gaffar Al-Ansari7 个月前

other model just pushing their agentic capabilities ... google push science you have my respect ...

Emerald Apple 的头像
Emerald Apple7 个月前

Looks like @grok has some catching up to do! I wonder what Grok 4.2 will fare in this match-up?

Vladimir 的头像
Vladimir7 个月前

everyone was watching the anthropic vs openai show and google just quietly posted the highest score on the board

Choblin 的头像
Choblin7 个月前

@Vicrom1509 Why are you not giving deep think model to pro users? 😞

tiangsae 的头像
tiangsae7 个月前

no for pro users? what's wrong with the AI world and its subscription plan right now?

Jonathan Ng 的头像
Jonathan Ng7 个月前

Gotta love these rate limits despite me paying already 250$ a month

Mahesh Malviya 的头像
Mahesh Malviya7 个月前

84.6% on ARC-AGI 2, and it's Only February of 2026... Google cooked...

eclectic 的头像
eclectic7 个月前

Google trains and runs Gemini (incl. Gemini 3 Pro) primarily on its own TPUs, not Nvidia GPUs. Why this matters: • TPUs = custom AI chips (ASICs) built in-house since 2013 • Higher efficiency, lower cost, and better scaling for Google’s workloads • Independence from Nvidia pricing + supply constraints • DeepMind co-designs TPUs → tight model–hardware optimization

JoshXT 的头像
JoshXT7 个月前

The new best is the new worst it will ever be. Looks like a great model!

mpantic3 的头像
mpantic37 个月前

The real frontier isn’t AI replacing scientists — it’s scientists expanding what they can attempt.

Ryan Swanson 的头像
Ryan Swanson7 个月前

its all making sense now

𝗿𝗮𝗺𝗮𝗸𝗿𝘂𝘀𝗵𝗻𝗮— 𝗲/𝗮𝗰𝗰 的头像
𝗿𝗮𝗺𝗮𝗸𝗿𝘂𝘀𝗵𝗻𝗮— 𝗲/𝗮𝗰𝗰7 个月前

Google is always going to win the AI race. Not even close.

Dacist Rapian 的头像
Dacist Rapian7 个月前

My cousin used Gemini to create a serum that would turn his eyes gray. He went blind.

Mayank 的头像
Mayank7 个月前

That’s not “AI writing essays.” That’s AI accelerating science.

CleanPegasus (🌎, 💻 ) 的头像
CleanPegasus (🌎, 💻 )7 个月前

How do I use deep think in the Gemini app?

Mahesh Malviya 的头像
Mahesh Malviya7 个月前

Other model just pushing their agentic capabilities... Google push Science. You have My Respect...

Conor Dart 的头像
Conor Dart7 个月前

Amazing progress in a few months, imagine what models we don't get to see, that they use internally!

Earl St Sauver 的头像
Earl St Sauver7 个月前

I’d love to see a frontiermath benchmark!

Waldrada 的头像
Waldrada7 个月前

"Please give Google your cutting edge tech designs" said no one sane, ever.

Lumenveil ✧ 的头像
Lumenveil ✧7 个月前

is the limit still 10 a day?

Jim Scott 的头像
Jim Scott7 个月前

AI that moves from theory to practical application is where the real disruption happens. Semiconductor design is just the beginning. I can't believe how much progress has been happening with these models.

小威Volt⚡️ 的头像
小威Volt⚡️7 个月前

The semiconductor materials design use case is a perfect demo — that's exactly where deep reasoning shines over pattern matching. 84.6% on ARC-AGI-2 is wild though, wasn't this benchmark supposed to be hard for years?

Restore The West 🇺🇸 的头像
Restore The West 🇺🇸7 个月前

Google and every single AI Company are going to cause Planetary Extinction

Arsalan Ahmed 的头像
Arsalan Ahmed7 个月前

This is amazing. As I said earlier, Ai wil start evolving

NexasTech 的头像
NexasTech7 个月前

Using AI to design semiconductors is the recursive loop everyone predicted but nobody's shipped at production scale. The real question is whether Deep Think compresses the 18-month fab iteration cycle. If Duke's lab goes from simulation to tape-out in weeks instead of months, that's the benchmark worth tracking.

Mstr DiVino 的头像
Mstr DiVino7 个月前

I don't see the reason to continue paying for the Gemini_Pro version. It is significantly inferior to its competitors at a similar price. You could offer 2-3 requests per day to Deep Think.

Apollo 的头像
Apollo7 个月前

does this concern internal Gemini Deep Think model not available to public?

相关视频