Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Introducing Drama 3 (preview): the most controllable TTS model ever. Describe the tone, pacing, and character in simple language, without thinking in audio tags. Shift your voice mid-sentence, generate a multi-character scene, or fix only a single word. Comment DRAMA to get an API key 🐟

155,993 Aufrufe • vor 3 Tagen •via X (Twitter)

45 Kommentare

Profilbild von Voyager1919
Voyager1919vor 3 Tagen

DRAMA! Will gladly test it in my new game! It's a FNAF-like and I need a phone guy :) Thanks!

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

so exciting! plz check your dm and can't wait to see your game!

Profilbild von Ravi Kushwaha
Ravi Kushwahavor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! 🪄

Profilbild von Yogsho
Yogshovor 3 Tagen

DRAMA

Profilbild von Ryu
Ryuvor 3 Tagen

DRAMA

Profilbild von Yoland Yan
Yoland Yanvor 3 Tagen

Congrats on the launch! DRAMA!!

Profilbild von 🐨코알라🐨
🐨코알라🐨vor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! ✨

Profilbild von Yu / ゆう
Yu / ゆうvor 3 Tagen

DRAMA

Profilbild von りょうま 🚀 AIで有意義な暮らしを創るエンジニア
りょうま 🚀 AIで有意義な暮らしを創るエンジニアvor 3 Tagen

DRAMA

Profilbild von kapio
kapiovor 3 Tagen

DRAMA

Profilbild von 角煮星丸
角煮星丸vor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! 🪄

Profilbild von tane@AI
tane@AIvor 3 Tagen

DRAMA

Profilbild von StudioYebisu
StudioYebisuvor 3 Tagen

Is my comment getting ignored? DRAMA! Please~!

Profilbild von 🥷
🥷vor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm📫

Profilbild von APTMA
APTMAvor 3 Tagen

DRAMA

Profilbild von LUTHANDO 💙
LUTHANDO 💙vor 3 Tagen

DRAMA

Profilbild von 裏方こいし
裏方こいしvor 3 Tagen

DRAMA

Profilbild von 氪学家
氪学家vor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! 🪄

Profilbild von どすを (DOSUO) | 📈PCパーツ価格動向📊 / AI・LLM
どすを (DOSUO) | 📈PCパーツ価格動向📊 / AI・LLMvor 3 Tagen

DRAMA

Profilbild von 空想圏域 / Imaginal Domain|編纂室
空想圏域 / Imaginal Domain|編纂室vor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! 🪄

Profilbild von ゆーき
ゆーきvor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! ✨

Profilbild von ほしりす
ほしりすvor 3 Tagen

DRAMA

Profilbild von あんり
あんりvor 3 Tagen

DRAMA

Profilbild von 残念院さん 貴方のファンアートを糧にすくすく育つ怪物系教祖
残念院さん 貴方のファンアートを糧にすくすく育つ怪物系教祖vor 3 Tagen

DRAMA

Profilbild von Ted Nova
Ted Novavor 3 Tagen

Drama

Profilbild von Wh1z
Wh1zvor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! 🪄

Profilbild von Nicku
Nickuvor 3 Tagen

DRAMA i love fish audio i've been making good use of the free month on the new model!

Profilbild von もっく
もっくvor 3 Tagen

DRAMA

Profilbild von Anderson
Andersonvor 3 Tagen

your voice is now open source

Profilbild von とよ@建ログ
とよ@建ログvor 3 Tagen

DRAMA

Profilbild von ラスワン|20億溶かした元上場社長|次は100億
ラスワン|20億溶かした元上場社長|次は100億vor 3 Tagen

DRAMA

Profilbild von シュン@G.B.'s Studio | OSAKA
シュン@G.B.'s Studio | OSAKAvor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! 🪄

Profilbild von route@LOVOTオーナー(AITuber切り抜きちゃん休止中?)
route@LOVOTオーナー(AITuber切り抜きちゃん休止中?)vor 3 Tagen

DRAMA

Profilbild von Fish Audio
Fish Audiovor 3 Tagen

check your dm! ✨

Profilbild von ともか
ともかvor 3 Tagen

DRAMA!

Profilbild von yukke
yukkevor 3 Tagen

DRAMA

Ähnliche Videos

VoxCPM 2 just dropped by OpenBMB Only 2B-param open-source TTS (Text-to-Speech) model built for production-grade multilingual voice work. Apache-2.0 license, Can run on only 8GB VRAM. • Eliminates the "robotic" feel of traditional TTS, delivering prosody and emotional depth suitable for high-stakes professional environments like filmmaking, gaming, animation, and audiobooks. • 30-language multilingual: no language tag needed, just type in a supported language and generate directly. • Voice design: create a brand-new voice from a text description alone, like age, tone, pace, or emotion. No reference audio required. Describe the desired voice characteristics (gender, age, tone, emotion, pace …) in Control Instruction, and VoxCPM2 will craft a unique voice from your description alone. • Controllable cloning: clone from a short clip, then steer delivery style without losing the speaker’s core voice. • Ultimate cloning: use reference audio + transcript for continuation-style cloning that keeps the tiny vocal details. • 48kHz output: takes 16kHz reference audio and produces studio-quality speech without an external upsampler. • Real-time ready: around 0.3 RTF on RTX 4090, even lower with Nano-VLLM. • Commercial use: Apache-2.0 licensed. Developer-Friendly Infrastructure: - Native Torch Inference: Direct support for PyTorch-based workflows. - Training Flexibility: Supports both full-parameter and LoRA fine-tuning for specific domain adaptation. - Production Readiness: Compatible with voxcpm-nanovllm for large-scale, high-concurrency deployment.

Rohan Paul

13,541 Aufrufe • vor 5 Monaten