Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Meet LongCat-Video-Avatar 1.5🐱—our upgraded, open-source digital human framework. Built for real production, not just short demos. What's New: 🔹 Upgraded Audio Encoder: Replaces Wav2Vec2 with Whisper-Large, yielding significantly smoother and more natural lip dynamics. 🔹 Production-Ready Stability: Achieves accurate lip-synchronization, full-body temporal stability, and robust long-video generation with strict...

31,524 görüntüleme • 3 ay önce •via X (Twitter)

11 Yorum

Garry profil fotoğrafı
Garry3 ay önce

How much time it will take to generate 60 seconds of avatar video (720p) with audio upload?

ZIAISTAN profil fotoğrafı
ZIAISTAN3 ay önce

Amazing

blankbrain profil fotoğrafı
blankbrain3 ay önce

oh finally , some good open sauce stuff from good plebs

Alice The Ai Expert profil fotoğrafı
Alice The Ai Expert3 ay önce

Open digital humans leveled up

T1000 profil fotoğrafı
T10003 ay önce

@toyxyz3 This is great.

Ruzaina profil fotoğrafı
Ruzaina3 ay önce

Open-source avatars are leveling up

. profil fotoğrafı
.3 ay önce

the question is: it is better then ltx 2.3? its faster? if dont, it is useless

𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏 profil fotoğrafı
𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏3 ay önce

whisper swap is the right call, wav2vec2 was always the weak link for lip sync. Curious what your real-time latency looks like end-to-end on a production load

Sani Ai Tech profil fotoğrafı
Sani Ai Tech3 ay önce

Open-source digital humans with 8-step inference and solid lip sync is a big deal

chener profil fotoğrafı
chener3 ay önce

@HeyGen

Artyom profil fotoğrafı
Artyom3 ay önce

Is this based on wan 2.2 video model?

Benzer Videolar