Loading video...

Video Failed to Load

Go Home

Today, we’re launching Gemini 3.5 Transcribe, our new speech-to-text model with sub-second streaming and intelligent post-processing for agent interfaces. 2.6% WER on non-streaming and 4.0% on streaming. Supports 85+ languages, cleans up conversational disfluencies ("um", "ah", and mid-sentence self-corrections), handles alphanumeric tokens (postal codes or IDs), 70% reduction time...

15,024 views • 5 days ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos