Loading video...

Video Failed to Load

Go Home

Soniox has a voice for every use case and style. Soniox TTS v2 delivers natural-sounding speech for storytelling, customer support, social media, education, voice agents, and many more. Adapt the delivery with audio tags for vocalizations, style, speed and emotion in 60+ languages. Explore Soniox voice library:

72,883 views • 24 days ago •via X (Twitter)

2 Comments

RAZA | AI EXPLORER's profile picture
RAZA | AI EXPLORER23 days ago

That’s an impressive creative pipeline. AI, 3D, and video working together with real control over the final result.

Mauro Mequelussi's profile picture
Mauro Mequelussi24 days ago

Is there any plan to make soundtracks or sound effects available to use alongside TTS?

Related Videos

VoxCPM 2 just dropped by OpenBMB Only 2B-param open-source TTS (Text-to-Speech) model built for production-grade multilingual voice work. Apache-2.0 license, Can run on only 8GB VRAM. • Eliminates the "robotic" feel of traditional TTS, delivering prosody and emotional depth suitable for high-stakes professional environments like filmmaking, gaming, animation, and audiobooks. • 30-language multilingual: no language tag needed, just type in a supported language and generate directly. • Voice design: create a brand-new voice from a text description alone, like age, tone, pace, or emotion. No reference audio required. Describe the desired voice characteristics (gender, age, tone, emotion, pace …) in Control Instruction, and VoxCPM2 will craft a unique voice from your description alone. • Controllable cloning: clone from a short clip, then steer delivery style without losing the speaker’s core voice. • Ultimate cloning: use reference audio + transcript for continuation-style cloning that keeps the tiny vocal details. • 48kHz output: takes 16kHz reference audio and produces studio-quality speech without an external upsampler. • Real-time ready: around 0.3 RTF on RTX 4090, even lower with Nano-VLLM. • Commercial use: Apache-2.0 licensed. Developer-Friendly Infrastructure: - Native Torch Inference: Direct support for PyTorch-based workflows. - Training Flexibility: Supports both full-parameter and LoRA fine-tuning for specific domain adaptation. - Production Readiness: Compatible with voxcpm-nanovllm for large-scale, high-concurrency deployment.

Rohan Paul

13,541 views • 5 months ago