Loading video...
Video Failed to Load
HOLY CRAP, a new super tiny 1.6B param voice model just dropped that seems to.. outperform 11labs!? 😵💫 From Nari-labs, Dia is an Apache 2.0 voice model, that can generate laughs, sniffs and emotions, copy an existing voice and is effectively real time on larger GPUs:
525,803 views • 1 year ago •via X (Twitter)
10 Comments

We'll of course cover this on the next @thursdai_pod (thanks for the heads up @rodrimora !)

Here's the links: Demo Page: Github: HF Model: Try It:

Coming to MLX Audio this week 🔥 @lllucas and I are on it 🚀

@lllucas can't wait to try this on my mac!

The best sniff emotion I have heard so far Other models feel like a sigh

(sniff) the laughter is so dope as well, it comes in juuust before and you can hear the smile on the invisible's AI face as it starts to render the laughter

Intonation and tone are great, but quality is nowhere near elevenlabs.

Honestly voice generation feels like it has been stuck at the same level for at least a year, not real improvement

@turing_hamster

This on Pinokio yet @cocktailpeanut ?
