Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

🔈 today we're introducing two new live dialogue models Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intelligence and multi-step reasoning. these audio models are available to build...

66,189 Aufrufe • vor 2 Tagen •via X (Twitter)

30 Kommentare

Profilbild von Alisa Fortin
Alisa Fortinvor 2 Tagen

We're so excited to bring these models to our customers! Please try and share your feedback

Profilbild von iliyas
iliyasvor 2 Tagen

Why the release data is 20 August 2026 in both models ??

Profilbild von Rama Puvvula
Rama Puvvulavor 2 Tagen

Sharing screen freezes the voice conversation.

Profilbild von Shaka_Fort
Shaka_Fortvor 2 Tagen

Could you make Gemini Live's voice sound more natural? It sounds very robotic compared to Chatgpt's voices.

Profilbild von GGUFzy
GGUFzyvor 2 Tagen

Trying the live dialogue models in Studio.

Profilbild von Apollo
Apollovor 2 Tagen

Google will release Gemini Live (just to pass Open AI on something) Gemini Flash Lite (for uses no one asked for) Gemini Cyber (because Dario is hyping cyberfears) before releasing Gemini Pro. which is a same because Gemini pro models used to be frontier tier

Profilbild von KC
KCvor 2 Tagen

中文支持如何

Profilbild von RAZA | AI EXPLORER
RAZA | AI EXPLORERvor 2 Tagen

Live dialogue plus visual grounding is a powerful combination.

Profilbild von Rokas Remeika | LiveKit, WebRTC, SIP/PSTN
Rokas Remeika | LiveKit, WebRTC, SIP/PSTNvor 2 Tagen

Great. I wonder how does this model perform in foreign language accuracy

Profilbild von בינה מלאכותית בעברית
בינה מלאכותית בעבריתvor 2 Tagen

Would be handy to compare both on the same conversation, including cost and response time.

Profilbild von TBK Joshua
TBK Joshuavor 2 Tagen

No frontier model from google yet... really looking for a model thats more efficient in 3d and computer use than astra... as a team gemini user, really hoping this wait is worth it...

Profilbild von Smoky Co
Smoky Covor 1 Tag

NOTICE : do not waste your time fucking with this. It doesn't even work. @GoogleAIStudio ; [Notice: gemini-3.8-flash is at high demand (503). Auto-falling back to gemini-3.5-flash...]

Profilbild von Fly
Flyvor 2 Tagen

This sounds terrible.

Profilbild von Nick Dry
Nick Dryvor 2 Tagen

I think the development or app creation section should be updated with (🟢🟡🔴)

Profilbild von Arthur.eth
Arthur.ethvor 1 Tag

Gemini的视觉理解力依然傲视群雄

Profilbild von indoor positioning
indoor positioningvor 2 Tagen

很赞

Profilbild von Immortus
Immortusvor 2 Tagen

is it a full duplex model?

Profilbild von Alex
Alexvor 2 Tagen

Gemini 3.8 Live just showed up in AI Studio live dialogue. going to burn a call on it tonight and see if latency feels different from the old live stack.

Profilbild von Armstrong Too | Photo + Film + AI
Armstrong Too | Photo + Film + AIvor 2 Tagen

The async function-calling detail is the real interface shift. A voice model that can acknowledge, work, and return with grounded context feels less like a turn-based chatbot and more like a task runner. Latency communication becomes part of trust.

Profilbild von Vrushali
Vrushalivor 2 Tagen

Gr8

Profilbild von APIMart Labs
APIMart Labsvor 2 Tagen

The Live vs Extended Thinking split is clearer here than in the consumer posts. One for cheap dialogue, one for calls that actually need to think mid-conversation.

Profilbild von lenny anzalichi
lenny anzalichivor 2 Tagen

Is Gemini 3.8 live extended thinking also more expensive at API cost then normal Gemini 3.8 live ?

Profilbild von AIBotics
AIBoticsvor 2 Tagen

I can finally build a voice assistant that understands my screen without draining my API credits in a single afternoon.

Profilbild von Ibesh
Ibeshvor 2 Tagen

the launch pitch says fluid dialogue, but the thread is already reporting freezes during screen sharing. reliability in the awkward edge cases will decide this product.

Profilbild von AI Mastery Guide
AI Mastery Guidevor 2 Tagen

two at once is huge

Profilbild von Omololu
Omololuvor 1 Tag

Google’s new voice models don’t pause to “think.” Gemini 3.8 Live reasons while it talks and runs tools in the background. Extended Thinking just took #1 on speech-to-speech. If voice agents finally work, which product dies first - phone trees or meeting notes?

Profilbild von linn*|lindreww*
linn*|lindreww*vor 2 Tagen

would love for high quality, cheap TTS. pleaaaassseeee

Profilbild von Gunnleygur Djurhuus
Gunnleygur Djurhuusvor 2 Tagen

Please make it work in Faroese. We really need a good transcription model to help handicapped people in our country.

Profilbild von Nexqor
Nexqorvor 2 Tagen

what's the interruption latency, that's the real live-audio metric

Profilbild von youfeng
youfengvor 2 Tagen

The split is the interesting bit. In a live audio model, extended thinking means the model goes quiet before it speaks — you spend latency to buy deliberation, and nothing gives you both. So the two SKUs are really one dial: how long will a user wait before a voice comes back?

Ähnliche Videos