Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Today we’re introducing Scribe v2: the most accurate transcription model ever released. While Scribe v2 Realtime is optimized for ultra low latency and agents use cases, Scribe v2 is built for batch transcription, subtitling, and captioning at scale.

558,058 Aufrufe • vor 9 Monaten •via X (Twitter)

40 Kommentare

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Scribe v2 achieves the lowest word error rate based on industry-standard benchmarks. Scribe v2 improves on the state of the art stability from Scribe v1. It handles pauses, changes in tone and delivery, together with long silences without any issues, delivering unmatched accuracy across more than 90 languages.

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Scribe v2 is now used in ElevenLabs Studio for more accurate subtitles, captions and transcriptions, supporting teams that manage large libraries of audio and video across marketing, media, research, training, and compliance use cases.

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Keyterm Prompting. Keyterm prompting goes beyond standard Custom Vocabulary by using the transcript’s context. Select up to 100 words or phrases, and Scribe v2 will accurately decide when to transcribe those terms.

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Entity Detection. Select up to 56 categories across Personally Identifiable Information, health data or payment details. Scribe v2 will automatically detect these instances and their exact timestamps in your transcript. Read the docs:

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Smart Multi-language Support. Send audio with multiple languages and Scribe v2 will automatically detect and transcribe in the right language.

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Other key features: - Smart Speaker Diarization: Intuitive labeling of every speaker for clear, organized transcripts - Precise Word-Level Timestamps: Capture the exact moment each word is spoken. Scribe v2 detailed timestamps enable seamless subtitle syncing and interactive experiences - Dynamic Audio Tagging: From laughter to footsteps, Scribe tags every sound event, enriching your transcripts with the full context of your audio - Enterprise ready with SOC 2, ISO27001, PCI DSS L1, HIPAA, GDPR compliance, EU & India residency and support for zero retention mode

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Build with the API. With Scribe v2, developers and enterprises can automate complex audio pipelines, achieve higher accuracy in global content workflows, and scale with full compliance and data residency controls. Read the docs:

Profilbild von ElevenLabs
ElevenLabsvor 9 Monaten

Try Scribe v2 today

Profilbild von Richard Reis
Richard Reisvor 9 Monaten

Scribe v3 feature request, please remove "ums" and "ahs" from transcripts 🙏

Profilbild von Guillermo Rauch
Guillermo Rauchvor 9 Monaten

cool

Profilbild von Luke Harries
Luke Harriesvor 9 Monaten

This model is a beast! It's the big version of Scribe v2 Realtime Very excited to see what you all build and create with it

Profilbild von AshutoshShrivastava
AshutoshShrivastavavor 9 Monaten

This is awesome.. 👏

Profilbild von Avais Aziz
Avais Azizvor 9 Monaten

Scribe v2 feels like the moment speech-to-text finally grew up. Crystal-clear accuracy across dozens of languages, thoughtful keyterm guidance, and speaker diarization that actually works. ElevenLabs just raised the bar again. Brilliant work.

Profilbild von Thais Castello Branco
Thais Castello Brancovor 9 Monaten

who made the video? great stuff!

Profilbild von Contextrix
Contextrixvor 9 Monaten

Scribe v2 being optimized for batch transcription at scale is perfect for subtitling workflows. The improved accuracy will make it much easier to handle large volumes of content reliably.

Profilbild von Happy yadav
Happy yadavvor 9 Monaten

Can it really beat @GroqInc ‘s Whisper Large v3 Turbo ? I feel no.

Profilbild von Subrahmanya Gaonkar
Subrahmanya Gaonkarvor 9 Monaten

I'm using @elevenlabs & @suno etc.. for building an open-source version of @goClueso. It turns raw screen recordings into polished product demos. check out!!

Profilbild von Micah Berkley - TheAIMogul
Micah Berkley - TheAIMogulvor 9 Monaten

Why do these companies always compare their model to these stupid old legacy models..... #DoingTooMuch

Profilbild von Umesh
Umeshvor 9 Monaten

So awesome! 🔥 ElevenLabs keeps raising the bar.

Profilbild von roshi
roshivor 9 Monaten

Why are you guys not letting me generate :( . Does Elevenlabs have a country blocklist?

Profilbild von Wesley
Wesleyvor 9 Monaten

Scribe v2 looks absolutely game-changing! The benchmarks are insane—lowest WER across 90+ languages, smart keyterm prompting, entity detection, and that enterprise-grade compliance. Can't wait to test it for long-form podcasts and multilingual subs. Huge congrats @elevenlabs team!

Profilbild von Abdulmuiz Adeyemo
Abdulmuiz Adeyemovor 9 Monaten

This matters more than people think. Accurate batch transcription = better datasets, better agents, better products. Realtime is cool. But scale + accuracy is where real businesses are built.

Profilbild von Morgan
Morganvor 9 Monaten

So awesome, can’t imagine how complex it is to get to this level of accuracy, congrats to the whole team! Can’t wait to play around with it.

Profilbild von Sakib
Sakibvor 9 Monaten

@superwhisper

Profilbild von Felicia_k_o
Felicia_k_ovor 9 Monaten

Hello. I live in Kazakhstan and can't pay for your service. It doesn't accept my bank card. This is a problem for all users I know in my country. Could you please tell me when this issue will be resolved?

Profilbild von Shamim Hossain
Shamim Hossainvor 9 Monaten

That's impressive

Profilbild von luizfcouto
luizfcoutovor 9 Monaten

@grok , what’s de diff between v2 and v2 realtime and give me wild examples of what we can accomplish today with this new version?

Profilbild von Ben Ganz
Ben Ganzvor 9 Monaten

@grok is it cheaper than OpenAI’s

Profilbild von Kashif Iqbal
Kashif Iqbalvor 9 Monaten

Elevanlab new method

Profilbild von Slumdog8
Slumdog8vor 9 Monaten

How fast is the async file processing via api?

Profilbild von Sa'id Adam
Sa'id Adamvor 9 Monaten

Fantastic

Profilbild von Narek
Narekvor 9 Monaten

Have you fixed Armenian? 👌

Profilbild von Himanshu Kumar
Himanshu Kumarvor 9 Monaten

@elevenlabs, the accuracy improvements in Scribe v2 will significantly help with large-scale subtitling projects.

Profilbild von 이상범
이상범vor 8 Monaten

i need to hallucinate break this model next

Profilbild von Jesse Jameson
Jesse Jamesonvor 9 Monaten

This would be great news if you were not exploiting voice creators.

Profilbild von rouven
rouvenvor 9 Monaten

i love the animations in this video. it sooo clean

Profilbild von Sasha
Sashavor 9 Monaten

What is the price per hour of transcription?

Profilbild von Fairooz Choudhury
Fairooz Choudhuryvor 9 Monaten

Exciting news! Improving accuracy is a game changer.

Profilbild von Rahul Vyas
Rahul Vyasvor 9 Monaten

@grok what does it cost to transcribe 80 hours of audio (approx 30 min each)

Profilbild von Gonzalo Cordova
Gonzalo Cordovavor 9 Monaten

don’t know about the model yet, but the release video is pretty cool, gotta give them that

Ähnliche Videos