正在加载视频...

视频加载失败

EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine github: EmotiVoice is a powerful and modern open-source text-to-speech engine. EmotiVoice speaks both English and Chinese, and with over 2000 different voices. The most prominent feature is emotional synthesis, allowing you to create speech with a wide range of emotions, including...

312,365 次观看 • 2 年前 •via X (Twitter)

10 条评论

Furkan Gözükara 的头像
Furkan Gözükara2 年前

Even demo is low sound quality

Andrzej Białecki 的头像
Andrzej Białecki2 年前

I wonder when we'll have singing voice synthesis guided by text and midi notes of a lead sound.

Jeff Araujo 的头像
Jeff Araujo2 年前

@camenduru, would be awesome to have a Colab available using this Engine 🥹

Fran Abenza 的头像
Fran Abenza2 年前

Would it run in M1, 8Gb Ram?

Nathan Odle 的头像
Nathan Odle2 年前

I tried running it locally and didn't get much variation between emotion prompts. Tried different (english) voices and happy/angry pretty much sounded the same most of the time. Maybe it works better with chinese?

Youdao Open Source 的头像
Youdao Open Source2 年前

Author here. Thanks for your interest in the project. We will post a roadmap for future updates shortly.

Patrick's AIBuzzNews 的头像
Patrick's AIBuzzNews2 年前

Does it outperform Bark?

Ai News 24/7 的头像
Ai News 24/72 年前

EmotiVoice sounds amazing, especially with its prompt-controlled feature. Gonna give it a try!

Ping Chen 的头像
Ping Chen2 年前

@Memdotai mem it

tinyfish 的头像
tinyfish2 年前

Should try

相关视频