Загрузка видео...

Не удалось загрузить видео

На главную

Meet LongCat-Video-Avatar 1.5🐱—our upgraded, open-source digital human framework. Built for real production, not just short demos. What's New: 🔹 Upgraded Audio Encoder: Replaces Wav2Vec2 with Whisper-Large, yielding significantly smoother and more natural lip dynamics. 🔹 Production-Ready Stability: Achieves accurate lip-synchronization, full-body temporal stability, and robust long-video generation with strict...

31,524 просмотров • 3 месяцев назад •via X (Twitter)

Комментарии: 11

Фото профиля Garry
Garry3 месяцев назад

How much time it will take to generate 60 seconds of avatar video (720p) with audio upload?

Фото профиля ZIAISTAN
ZIAISTAN3 месяцев назад

Amazing

Фото профиля blankbrain
blankbrain3 месяцев назад

oh finally , some good open sauce stuff from good plebs

Фото профиля Alice The Ai Expert
Alice The Ai Expert3 месяцев назад

Open digital humans leveled up

Фото профиля T1000
T10003 месяцев назад

@toyxyz3 This is great.

Фото профиля Ruzaina
Ruzaina3 месяцев назад

Open-source avatars are leveling up

Фото профиля .
.3 месяцев назад

the question is: it is better then ltx 2.3? its faster? if dont, it is useless

Фото профиля 𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏
𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏3 месяцев назад

whisper swap is the right call, wav2vec2 was always the weak link for lip sync. Curious what your real-time latency looks like end-to-end on a production load

Фото профиля Sani Ai Tech
Sani Ai Tech3 месяцев назад

Open-source digital humans with 8-step inference and solid lip sync is a big deal

Фото профиля chener
chener3 месяцев назад

@HeyGen

Фото профиля Artyom
Artyom3 месяцев назад

Is this based on wan 2.2 video model?

Похожие видео