正在加载视频...

视频加载失败

Here are the best practices for using Eleven v3 (alpha) - the most expressive Text to Speech model.

43,851 次观看 • 1 年前 •via X (Twitter)

11 条评论

ElevenLabs 的头像
ElevenLabs1 年前

1. Use longer prompts. Eleven v3 performs better with longer inputs. Prompts shorter than 250 characters are more likely to produce unstable results.

ElevenLabs 的头像
ElevenLabs1 年前

2. Pick the right voice. Some voices are higher quality and more expressive than others. Use voices made for the language you're working in. When creating new voices, include a wider emotional range than before. Explore 22 voices that perform well with v3:

ElevenLabs 的头像
ElevenLabs1 年前

3. Use audio tags to control delivery. Audio tags like [sarcastic], [whispers], [excited], or [strong French accent] shape how the model speaks. Choose a voice suited to your intended delivery. Don’t expect a whispering voice to shout convincingly. Audio tags aren’t universal.

ElevenLabs 的头像
ElevenLabs1 年前

4. Keep experimenting. Eleven v3 (alpha) is a research preview. It often requires more prompt engineering than earlier models, but the results are breathtaking.

ElevenLabs 的头像
ElevenLabs1 年前

Read the full best practices guide:

AssemblyAI 的头像
AssemblyAI1 年前

Our speech-to-text models are the most accurate on the market with top rankings across industry benchmarks. - The highest accuracy rates—up to 95% - Up to 30% fewer hallucinations than other leaders - Low latency—63 minutes converts in 35 seconds Try via API for free today 👇

Luke Harries 的头像
Luke Harries1 年前

Great explanation by @alecwilcock_

James McAulay ❙❙ ElevenLabs 的头像
James McAulay ❙❙ ElevenLabs1 年前

👏 @alecwilcock_

Lise Slimane 的头像
Lise Slimane1 年前

@alecwilcock_ dropping wisdom once again 📝🤓

𝓘𝓼𝓷'𝓽 𝓲𝓽 𝓲𝓻𝓸𝓷𝓲𝓬 ® 的头像
𝓘𝓼𝓷'𝓽 𝓲𝓽 𝓲𝓻𝓸𝓷𝓲𝓬 ®1 年前

nice

Mememuncher 的头像
Mememuncher1 年前

Thanks for the explainer. This is one of the coolest updates right now that I’m excited about. Will be playing with this as much as I can 💚 🙌

相关视频