Video yükleniyor...
Video Yüklenemedi
Meet LongCat-Video-Avatar 1.5🐱—our upgraded, open-source digital human framework. Built for real production, not just short demos. What's New: 🔹 Upgraded Audio Encoder: Replaces Wav2Vec2 with Whisper-Large, yielding significantly smoother and more natural lip dynamics. 🔹 Production-Ready Stability: Achieves accurate lip-synchronization, full-body temporal stability, and robust long-video generation with strict... show more
31,524 görüntüleme • 3 ay önce •via X (Twitter)
11 Yorum

Garry3 ay önce
How much time it will take to generate 60 seconds of avatar video (720p) with audio upload?

ZIAISTAN3 ay önce
Amazing

blankbrain3 ay önce
oh finally , some good open sauce stuff from good plebs

Alice The Ai Expert3 ay önce
Open digital humans leveled up

T10003 ay önce
@toyxyz3 This is great.

Ruzaina3 ay önce
Open-source avatars are leveling up

.3 ay önce
the question is: it is better then ltx 2.3? its faster? if dont, it is useless

𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏3 ay önce
whisper swap is the right call, wav2vec2 was always the weak link for lip sync. Curious what your real-time latency looks like end-to-end on a production load

Sani Ai Tech3 ay önce
Open-source digital humans with 8-step inference and solid lip sync is a big deal

chener3 ay önce
@HeyGen

Artyom3 ay önce
Is this based on wan 2.2 video model?

