Loading video...
Video Failed to Load
Soprano: An instant, ultra-lightweight TTS model for realistic speech; generates 10 hours of 32kHz audio in <20s; streams with <15ms latency using just 80M params & <1GB VRAM. Has some limitations and drawbacks.
111,542 views • 7 months ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
