正在加载视频...
视频加载失败
Big moment for text-to-speech. Qwen just open-sourced a text-to-speech model that lets you clone voices, design new ones, and control speech using natural language. Let me explain what I mean: You can literally tell it "speak in a cheerful tone with slight nervousness," and it actually does that. No... show more
0 条评论
暂无评论
原始帖子的评论将显示在这里
相关视频
0:16
Sensitive content
⚠️ NSFW ⚠️ Looks like Hume made a virtually uncensored voice-native LLM called Octave—the first language model built specifically for TTS. You can generate any type of voice you want with a prompt, and their WebUI tool can auto-generate corresponding example dialogue! The model grasps user intent from text quite well, but as you’ll hear in the second example there can be (somewhat unsettling) glitches inb4 glitch orgasms become a new fetish category 🙃
Pliny the Liberator 🐉󠅫󠄼󠄿󠅆󠄵󠄐󠅀󠄼󠄹󠄾󠅉󠅭
36,379 次观看 • 1 年前
