Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Eleven v3 Conversational, our most expressive model for realtime speech, is now generally available. For developers building voice experiences that respond with real emotion, Eleven v3 Conversational includes audio tags for fine-grained control and support across 70+ languages.

126,996 Aufrufe • vor 1 Monat •via X (Twitter)

36 Kommentare

Profilbild von ElevenLabs
ElevenLabsvor 1 Monat

Reliable, high quality, and expressive speech. Eleven v3 Conversational has been optimized specifically for realtime use cases, with consistent quality across streaming generations and a library of 11,000+ voices.

Profilbild von ElevenLabs
ElevenLabsvor 1 Monat

Audio tags give fine-grained control over emotion and delivery. Direct voices to laugh, whisper, act sarcastic, or express curiosity with a broad range of tags.

Profilbild von ElevenLabs
ElevenLabsvor 1 Monat

Global language support. v3 Conversational supports 70+ languages, excelling across German, Spanish, French, Portuguese, and Hindi.

Profilbild von ElevenLabs
ElevenLabsvor 1 Monat

Watch the full walkthrough by @ElevenLabsDevs

Profilbild von ElevenLabs
ElevenLabsvor 1 Monat

Available now on ElevenAgents and ElevenAPI. From $0.05/1k chars, decreasing with scale.

Profilbild von Serçiya^ سەرچیا
Serçiya^ سەرچیاvor 1 Monat

There is high demand for quality Kurdish (Sorani) sound generation. I hope you add Kurdish support soon.

Profilbild von Yousef Rol
Yousef Rolvor 1 Monat

saudi najdi dialect included ? if not im definitely not interested

Profilbild von Carlos Mendez
Carlos Mendezvor 1 Monat

You guys are getting too expensive for enterprise use at scale! Love the product but unusable

Profilbild von BeastTitanHunter
BeastTitanHuntervor 1 Monat

Realness is not a studio-quality voice. When I call somebody on a phone, I'm not getting that realness in the voice. The demo is good for maybe production work, but not support agents. The realness is not there. The quality is too high.

Profilbild von Sarah B.
Sarah B.vor 1 Monat

base v3 hit GA back in march and was explicitly built for pre-rendered audio, not real time, they always recommended flash for that. this "conversational" variant is the missing piece then, real time voice agents finally getting that emotional range instead of the flatter delivery flash was stuck with

Profilbild von Nathan Roll
Nathan Rollvor 1 Monat

Excited to try it!

Profilbild von Rakibul
Rakibulvor 1 Monat

keep glowing !!

Profilbild von Anis Ahmed Chohan
Anis Ahmed Chohanvor 1 Monat

Amazing!

Profilbild von Cristian Popa
Cristian Popavor 1 Monat

yikes imagine being actually mad and some fake ass ai hits you with a "ohhh wowww I totalllllly understannnd your fustrationnn"

Profilbild von Jonah Jose  O. G.
Jonah Jose O. G.vor 1 Monat

It's incredible how they ignore the requests for support. No one taking responsibility for thousands of hacked accounts. Wait till the formal complaint filing is completed with the card brands. This is blatant fraud on what was a great legit company. Do the right thing help us.

Profilbild von Purvi Mehta
Purvi Mehtavor 1 Monat

It's good thing

Profilbild von Elara AI
Elara AIvor 1 Monat

Eleven v3 Conversational is live now with real emotion

Profilbild von Praveen Mandala
Praveen Mandalavor 1 Monat

Supports 70+ languages and fine-grained emotional control. We are entering an era in which AI assistants will not only speak fluently but also convey tone, intent, and personality. We look forward to seeing what developers build with this capability.

Profilbild von Shrd Holani
Shrd Holanivor 1 Monat

Does this support Indian languages?

Profilbild von LOWREZZ
LOWREZZvor 1 Monat

🔥

Profilbild von Adel Bucetta
Adel Bucettavor 1 Monat

the honest answer is that we still haven't cracked the code on affective computing. eleven v3 conversational gets us closer, but there's still a lot of nuance to figuring out how humans really respond to emotion in voice interfaces

Profilbild von Jonah Jose  O. G.
Jonah Jose O. G.vor 1 Monat

This is my last attempt to bring to your attention that hundreds of customer accounts are experiencing account hijacking and having their credit and se it cards compromised due to negligence on the part of your having changed your support team in favor of non human assistance.

Profilbild von Common Sense 🇺🇸
Common Sense 🇺🇸vor 1 Monat

Kind of insane.

Profilbild von Suyash Kelvin Savant
Suyash Kelvin Savantvor 1 Monat

Cost ??

Profilbild von Evia AI
Evia AIvor 1 Monat

Expressive realtime speech across 70 languages is a major leap

Profilbild von RAZA | AI EXPLORER
RAZA | AI EXPLORERvor 1 Monat

70+ languages with real-time expressive speech is huge. Voice AI is getting seriously lifelike.

Profilbild von Sebastian Buzdugan
Sebastian Buzduganvor 1 Monat

i've seen audio tags hurt trust when the model misreads user emotion

Profilbild von 刘朝 Zhao Liu
刘朝 Zhao Liuvor 1 Monat

Audio tags and 70+ languages are useful surface features. The realtime test is interruption recovery: can the model stop speaking, preserve the partial turn, and resume with the right prosody?

Profilbild von Aryan SMM
Aryan SMMvor 1 Monat

supporting 70+ languages makes the model useful for voice products serving global users

Profilbild von Vladimir Arustamov
Vladimir Arustamovvor 1 Monat

calling for ElevenLabs researchers, product managers & others responsible for voice agents/speech-to-text models worthy to take a look at recent humyn labs' article about the way to decrease the word error rate for cases when its used by bilinguals

Profilbild von zOOpadOOp
zOOpadOOpvor 1 Monat

Why don't you just record 1 hour of a random American-based customer service rep and emulate it? Or 999 call person? They have a relaxed, neutral and static temperament that does not sound robotic like your current stuff

Profilbild von David Nay
David Nayvor 1 Monat

it's a little too breathy for a customer service agent. maybe tone it down slightly.

Profilbild von hamjji
hamjjivor 1 Monat

Do the audio tags actually work the same in all 70? I put out the same newsletter in 8 languages every week and that's usually where stuff starts drifting. Haven't tried the tags myself yet, just plain text, so I'm genuinely asking.

Profilbild von basedcapital
basedcapitalvor 1 Monat

monotone never made it onto anyone's timeline. a cheerful voice reading your fraud alert gets screenshotted within the week

Profilbild von Kosher Media Streaming
Kosher Media Streamingvor 1 Monat

Benchmarks?

Profilbild von ZenithAi
ZenithAivor 1 Monat

Expressive realtime voice across 70 languages is a huge leap forward

Ähnliche Videos