Загрузка видео...
Не удалось загрузить видео
Meet LongCat-Video-Avatar 1.5🐱—our upgraded, open-source digital human framework. Built for real production, not just short demos. What's New: 🔹 Upgraded Audio Encoder: Replaces Wav2Vec2 with Whisper-Large, yielding significantly smoother and more natural lip dynamics. 🔹 Production-Ready Stability: Achieves accurate lip-synchronization, full-body temporal stability, and robust long-video generation with strict... show more
31,524 просмотров • 3 месяцев назад •via X (Twitter)
Комментарии: 11

Garry3 месяцев назад
How much time it will take to generate 60 seconds of avatar video (720p) with audio upload?

ZIAISTAN3 месяцев назад
Amazing

blankbrain3 месяцев назад
oh finally , some good open sauce stuff from good plebs

Alice The Ai Expert3 месяцев назад
Open digital humans leveled up

T10003 месяцев назад
@toyxyz3 This is great.

Ruzaina3 месяцев назад
Open-source avatars are leveling up

.3 месяцев назад
the question is: it is better then ltx 2.3? its faster? if dont, it is useless

𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏3 месяцев назад
whisper swap is the right call, wav2vec2 was always the weak link for lip sync. Curious what your real-time latency looks like end-to-end on a production load

Sani Ai Tech3 месяцев назад
Open-source digital humans with 8-step inference and solid lip sync is a big deal

chener3 месяцев назад
@HeyGen

Artyom3 месяцев назад
Is this based on wan 2.2 video model?

