Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

🔈 today we're introducing two new live dialogue models Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intelligence and multi-step reasoning. these audio models are available to build...

68,914 görüntüleme • 7 gün önce •via X (Twitter)

30 Yorum

Alisa Fortin profil fotoğrafı
Alisa Fortin7 gün önce

We're so excited to bring these models to our customers! Please try and share your feedback

iliyas profil fotoğrafı
iliyas7 gün önce

Why the release data is 20 August 2026 in both models ??

Rama Puvvula profil fotoğrafı
Rama Puvvula7 gün önce

Sharing screen freezes the voice conversation.

Shaka_Fort profil fotoğrafı
Shaka_Fort7 gün önce

Could you make Gemini Live's voice sound more natural? It sounds very robotic compared to Chatgpt's voices.

GGUFzy profil fotoğrafı
GGUFzy7 gün önce

Trying the live dialogue models in Studio.

Apollo profil fotoğrafı
Apollo7 gün önce

Google will release Gemini Live (just to pass Open AI on something) Gemini Flash Lite (for uses no one asked for) Gemini Cyber (because Dario is hyping cyberfears) before releasing Gemini Pro. which is a same because Gemini pro models used to be frontier tier

KC profil fotoğrafı
KC7 gün önce

中文支持如何

RAZA | AI EXPLORER profil fotoğrafı
RAZA | AI EXPLORER7 gün önce

Live dialogue plus visual grounding is a powerful combination.

Rokas Remeika | LiveKit, WebRTC, SIP/PSTN profil fotoğrafı
Rokas Remeika | LiveKit, WebRTC, SIP/PSTN7 gün önce

Great. I wonder how does this model perform in foreign language accuracy

בינה מלאכותית בעברית profil fotoğrafı
בינה מלאכותית בעברית7 gün önce

Would be handy to compare both on the same conversation, including cost and response time.

TBK Joshua profil fotoğrafı
TBK Joshua7 gün önce

No frontier model from google yet... really looking for a model thats more efficient in 3d and computer use than astra... as a team gemini user, really hoping this wait is worth it...

Smoky Co profil fotoğrafı
Smoky Co6 gün önce

NOTICE : do not waste your time fucking with this. It doesn't even work. @GoogleAIStudio ; [Notice: gemini-3.8-flash is at high demand (503). Auto-falling back to gemini-3.5-flash...]

Fly profil fotoğrafı
Fly6 gün önce

This sounds terrible.

Nick Dry profil fotoğrafı
Nick Dry6 gün önce

I think the development or app creation section should be updated with (🟢🟡🔴)

Arthur.eth profil fotoğrafı
Arthur.eth6 gün önce

Gemini的视觉理解力依然傲视群雄

indoor positioning profil fotoğrafı
indoor positioning7 gün önce

很赞

Immortus profil fotoğrafı
Immortus7 gün önce

is it a full duplex model?

Alex profil fotoğrafı
Alex7 gün önce

Gemini 3.8 Live just showed up in AI Studio live dialogue. going to burn a call on it tonight and see if latency feels different from the old live stack.

Armstrong Too | Photo + Film + AI profil fotoğrafı
Armstrong Too | Photo + Film + AI7 gün önce

The async function-calling detail is the real interface shift. A voice model that can acknowledge, work, and return with grounded context feels less like a turn-based chatbot and more like a task runner. Latency communication becomes part of trust.

Vrushali profil fotoğrafı
Vrushali7 gün önce

Gr8

APIMart Labs profil fotoğrafı
APIMart Labs6 gün önce

The Live vs Extended Thinking split is clearer here than in the consumer posts. One for cheap dialogue, one for calls that actually need to think mid-conversation.

lenny anzalichi profil fotoğrafı
lenny anzalichi7 gün önce

Is Gemini 3.8 live extended thinking also more expensive at API cost then normal Gemini 3.8 live ?

AIBotics profil fotoğrafı
AIBotics7 gün önce

I can finally build a voice assistant that understands my screen without draining my API credits in a single afternoon.

Ibesh profil fotoğrafı
Ibesh7 gün önce

the launch pitch says fluid dialogue, but the thread is already reporting freezes during screen sharing. reliability in the awkward edge cases will decide this product.

AI Mastery Guide profil fotoğrafı
AI Mastery Guide7 gün önce

two at once is huge

Omololu profil fotoğrafı
Omololu6 gün önce

Google’s new voice models don’t pause to “think.” Gemini 3.8 Live reasons while it talks and runs tools in the background. Extended Thinking just took #1 on speech-to-speech. If voice agents finally work, which product dies first - phone trees or meeting notes?

linn*|lindreww* profil fotoğrafı
linn*|lindreww*7 gün önce

would love for high quality, cheap TTS. pleaaaassseeee

Gunnleygur Djurhuus profil fotoğrafı
Gunnleygur Djurhuus7 gün önce

Please make it work in Faroese. We really need a good transcription model to help handicapped people in our country.

Nexqor profil fotoğrafı
Nexqor7 gün önce

what's the interruption latency, that's the real live-audio metric

youfeng profil fotoğrafı
youfeng7 gün önce

The split is the interesting bit. In a live audio model, extended thinking means the model goes quiet before it speaks — you spend latency to buy deliberation, and nothing gives you both. So the two SKUs are really one dial: how long will a user wait before a voice comes back?

Benzer Videolar