Загрузка видео...

Не удалось загрузить видео

На главную

🔈 today we're introducing two new live dialogue models Gemini 3.8 Live: built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: built for high-complexity tasks, with increased intelligence and multi-step reasoning. these audio models are available to build...

68,638 просмотров • 7 дней назад •via X (Twitter)

Комментарии: 30

Фото профиля Alisa Fortin
Alisa Fortin7 дней назад

We're so excited to bring these models to our customers! Please try and share your feedback

Фото профиля iliyas
iliyas7 дней назад

Why the release data is 20 August 2026 in both models ??

Фото профиля Rama Puvvula
Rama Puvvula6 дней назад

Sharing screen freezes the voice conversation.

Фото профиля Shaka_Fort
Shaka_Fort7 дней назад

Could you make Gemini Live's voice sound more natural? It sounds very robotic compared to Chatgpt's voices.

Фото профиля GGUFzy
GGUFzy6 дней назад

Trying the live dialogue models in Studio.

Фото профиля Apollo
Apollo7 дней назад

Google will release Gemini Live (just to pass Open AI on something) Gemini Flash Lite (for uses no one asked for) Gemini Cyber (because Dario is hyping cyberfears) before releasing Gemini Pro. which is a same because Gemini pro models used to be frontier tier

Фото профиля KC
KC7 дней назад

中文支持如何

Фото профиля RAZA | AI EXPLORER
RAZA | AI EXPLORER6 дней назад

Live dialogue plus visual grounding is a powerful combination.

Фото профиля Rokas Remeika | LiveKit, WebRTC, SIP/PSTN
Rokas Remeika | LiveKit, WebRTC, SIP/PSTN6 дней назад

Great. I wonder how does this model perform in foreign language accuracy

Фото профиля בינה מלאכותית בעברית
בינה מלאכותית בעברית6 дней назад

Would be handy to compare both on the same conversation, including cost and response time.

Фото профиля TBK Joshua
TBK Joshua7 дней назад

No frontier model from google yet... really looking for a model thats more efficient in 3d and computer use than astra... as a team gemini user, really hoping this wait is worth it...

Фото профиля Smoky Co
Smoky Co6 дней назад

NOTICE : do not waste your time fucking with this. It doesn't even work. @GoogleAIStudio ; [Notice: gemini-3.8-flash is at high demand (503). Auto-falling back to gemini-3.5-flash...]

Фото профиля Fly
Fly6 дней назад

This sounds terrible.

Фото профиля Nick Dry
Nick Dry6 дней назад

I think the development or app creation section should be updated with (🟢🟡🔴)

Фото профиля Arthur.eth
Arthur.eth6 дней назад

Gemini的视觉理解力依然傲视群雄

Фото профиля indoor positioning
indoor positioning6 дней назад

很赞

Фото профиля Immortus
Immortus7 дней назад

is it a full duplex model?

Фото профиля Alex
Alex7 дней назад

Gemini 3.8 Live just showed up in AI Studio live dialogue. going to burn a call on it tonight and see if latency feels different from the old live stack.

Фото профиля Armstrong Too | Photo + Film + AI
Armstrong Too | Photo + Film + AI6 дней назад

The async function-calling detail is the real interface shift. A voice model that can acknowledge, work, and return with grounded context feels less like a turn-based chatbot and more like a task runner. Latency communication becomes part of trust.

Фото профиля Vrushali
Vrushali7 дней назад

Gr8

Фото профиля APIMart Labs
APIMart Labs6 дней назад

The Live vs Extended Thinking split is clearer here than in the consumer posts. One for cheap dialogue, one for calls that actually need to think mid-conversation.

Фото профиля lenny anzalichi
lenny anzalichi6 дней назад

Is Gemini 3.8 live extended thinking also more expensive at API cost then normal Gemini 3.8 live ?

Фото профиля AIBotics
AIBotics6 дней назад

I can finally build a voice assistant that understands my screen without draining my API credits in a single afternoon.

Фото профиля Ibesh
Ibesh6 дней назад

the launch pitch says fluid dialogue, but the thread is already reporting freezes during screen sharing. reliability in the awkward edge cases will decide this product.

Фото профиля AI Mastery Guide
AI Mastery Guide6 дней назад

two at once is huge

Фото профиля Omololu
Omololu6 дней назад

Google’s new voice models don’t pause to “think.” Gemini 3.8 Live reasons while it talks and runs tools in the background. Extended Thinking just took #1 on speech-to-speech. If voice agents finally work, which product dies first - phone trees or meeting notes?

Фото профиля linn*|lindreww*
linn*|lindreww*6 дней назад

would love for high quality, cheap TTS. pleaaaassseeee

Фото профиля Gunnleygur Djurhuus
Gunnleygur Djurhuus6 дней назад

Please make it work in Faroese. We really need a good transcription model to help handicapped people in our country.

Фото профиля Nexqor
Nexqor6 дней назад

what's the interruption latency, that's the real live-audio metric

Фото профиля youfeng
youfeng7 дней назад

The split is the interesting bit. In a live audio model, extended thinking means the model goes quiet before it speaks — you spend latency to buy deliberation, and nothing gives you both. So the two SKUs are really one dial: how long will a user wait before a voice comes back?

Похожие видео