Загрузка видео...

Не удалось загрузить видео

На главную

🔈 today we're introducing two new live dialogue models Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intelligence and multi-step reasoning. these audio models are available to build...

65,207 просмотров • 1 день назад •via X (Twitter)

Комментарии: 30

Фото профиля Alisa Fortin
Alisa Fortin1 день назад

We're so excited to bring these models to our customers! Please try and share your feedback

Фото профиля iliyas
iliyas1 день назад

Why the release data is 20 August 2026 in both models ??

Фото профиля Rama Puvvula
Rama Puvvula1 день назад

Sharing screen freezes the voice conversation.

Фото профиля Shaka_Fort
Shaka_Fort1 день назад

Could you make Gemini Live's voice sound more natural? It sounds very robotic compared to Chatgpt's voices.

Фото профиля GGUFzy
GGUFzy1 день назад

Trying the live dialogue models in Studio.

Фото профиля Apollo
Apollo1 день назад

Google will release Gemini Live (just to pass Open AI on something) Gemini Flash Lite (for uses no one asked for) Gemini Cyber (because Dario is hyping cyberfears) before releasing Gemini Pro. which is a same because Gemini pro models used to be frontier tier

Фото профиля KC
KC1 день назад

中文支持如何

Фото профиля RAZA | AI EXPLORER
RAZA | AI EXPLORER1 день назад

Live dialogue plus visual grounding is a powerful combination.

Фото профиля Rokas Remeika | LiveKit, WebRTC, SIP/PSTN
Rokas Remeika | LiveKit, WebRTC, SIP/PSTN1 день назад

Great. I wonder how does this model perform in foreign language accuracy

Фото профиля בינה מלאכותית בעברית
בינה מלאכותית בעברית1 день назад

Would be handy to compare both on the same conversation, including cost and response time.

Фото профиля TBK Joshua
TBK Joshua1 день назад

No frontier model from google yet... really looking for a model thats more efficient in 3d and computer use than astra... as a team gemini user, really hoping this wait is worth it...

Фото профиля Smoky Co
Smoky Co1 день назад

NOTICE : do not waste your time fucking with this. It doesn't even work. @GoogleAIStudio ; [Notice: gemini-3.8-flash is at high demand (503). Auto-falling back to gemini-3.5-flash...]

Фото профиля Fly
Fly1 день назад

This sounds terrible.

Фото профиля Nick Dry
Nick Dry1 день назад

I think the development or app creation section should be updated with (🟢🟡🔴)

Фото профиля Arthur.eth
Arthur.eth1 день назад

Gemini的视觉理解力依然傲视群雄

Фото профиля indoor positioning
indoor positioning1 день назад

很赞

Фото профиля Immortus
Immortus1 день назад

is it a full duplex model?

Фото профиля Alex
Alex1 день назад

Gemini 3.8 Live just showed up in AI Studio live dialogue. going to burn a call on it tonight and see if latency feels different from the old live stack.

Фото профиля Armstrong Too | Photo + Film + AI
Armstrong Too | Photo + Film + AI1 день назад

The async function-calling detail is the real interface shift. A voice model that can acknowledge, work, and return with grounded context feels less like a turn-based chatbot and more like a task runner. Latency communication becomes part of trust.

Фото профиля Vrushali
Vrushali1 день назад

Gr8

Фото профиля APIMart Labs
APIMart Labs1 день назад

The Live vs Extended Thinking split is clearer here than in the consumer posts. One for cheap dialogue, one for calls that actually need to think mid-conversation.

Фото профиля lenny anzalichi
lenny anzalichi1 день назад

Is Gemini 3.8 live extended thinking also more expensive at API cost then normal Gemini 3.8 live ?

Фото профиля AIBotics
AIBotics1 день назад

I can finally build a voice assistant that understands my screen without draining my API credits in a single afternoon.

Фото профиля Ibesh
Ibesh1 день назад

the launch pitch says fluid dialogue, but the thread is already reporting freezes during screen sharing. reliability in the awkward edge cases will decide this product.

Фото профиля AI Mastery Guide
AI Mastery Guide1 день назад

two at once is huge

Фото профиля Omololu
Omololu1 день назад

Google’s new voice models don’t pause to “think.” Gemini 3.8 Live reasons while it talks and runs tools in the background. Extended Thinking just took #1 on speech-to-speech. If voice agents finally work, which product dies first - phone trees or meeting notes?

Фото профиля linn*|lindreww*
linn*|lindreww*1 день назад

would love for high quality, cheap TTS. pleaaaassseeee

Фото профиля Gunnleygur Djurhuus
Gunnleygur Djurhuus1 день назад

Please make it work in Faroese. We really need a good transcription model to help handicapped people in our country.

Фото профиля Nexqor
Nexqor1 день назад

what's the interruption latency, that's the real live-audio metric

Фото профиля youfeng
youfeng1 день назад

The split is the interesting bit. In a live audio model, extended thinking means the model goes quiet before it speaks — you spend latency to buy deliberation, and nothing gives you both. So the two SKUs are really one dial: how long will a user wait before a voice comes back?

Похожие видео