Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

New drop: Decart Lip Sync API. Real-time lip sync for any avatar. Stream audio → get perfectly synced video frames with low latency. No uncanny lag. No pre-rendering. Just living pixels that move when your model speaks. 🧵

43,663 görüntüleme • 10 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

I experimented a lot with the prompt until I landed on this version😎 Only face reference + exact 15-second audio. Generated in Higgsfield AI 🧩 Seedance 2.0. Small tip: when uploading reference audio, cut it precisely to whole seconds (6, 8, 10, 15…) with no leftover tails 🎵 Prompt: A cinematic 15-second music video shot in a dark modern dance studio with smooth grey reflective floor, black walls and horizontal neon light tubes. Continuous dynamic camera movement, mostly medium shots and close-ups, never too wide. The woman is constantly moving, no freezes or static poses. Main subject: young woman with messy medium-length wavy brown hair with bangs partially covering her face, freckles on nose and cheeks, blue-grey eyes, full glossy lips. She wears a tight black fishnet bodysuit. She is singing the entire time with clear, precise lip-sync, mouth actively moving, intense emotional expression. Lyrics: "Ye! Ye! You keep a box of names / In the drawer by your bed / Polaroids and ticket stubs / Stuffed under the red / You never take one thing / You take the whole last spark / Leave a little thumbprint / On every private heart". 0-2s: Tight medium close-up. She leans her upper body back, head tilting, singing passionately with strong lip-sync, hair falling over her face, body arching, one hand sliding across her chest. 2-4s: Camera slowly pushes in and circles. She comes out of the deep arch, torso still bent forward, hands on her thighs, lifting her head and looking straight into camera while singing with aggressive lip-sync. Three male dancers of different appearances (different ethnicities, hair styles and builds) wearing black tank tops and black wide pants are already close around her, moving with her in low tense postures. The other two men are visible at the edges of the frame, approaching. 4-7s: Medium close-up. She drops lower, body still in constant motion, hair swinging, singing intensely with clear mouth movements, sharp head turns, eyes locked on camera. Male dancers stay close, their hands lightly touching her as they move together. 8-12s: Medium shot with slow camera drift. Exactly five male dancers of completely different appearances, all wearing black tank tops and black wide pants, surround her tightly on the floor in a dense, intertwined formation. She is in the center, body still moving, upper body rising and shifting, singing with strong lip-sync. All five men move subtly with her, never static. The men never fully obscure her body. 13-15s: Dynamic medium shot. The five diverse male dancers lift her into the air in a powerful deep backbend. Her body is fully extended and arched, head thrown back, still singing with clear lip-sync. While holding her they gently rock her up and down in time with the beat. The camera starts from a clear side view of her arched body and smoothly transitions to a frontal view of her face. At the end they lower her smoothly onto her feet; she lands and immediately continues singing as the five men stay low on the floor around her. The men never fully obscure her. High fashion dance energy, sweaty skin, sharp timing, continuous fluid motion of the woman, priority on accurate lip-sync in every frame. Rules: no hand morphing, no body distortions, clean stable anatomy, fingers and hands remain consistent and natural throughout the entire video. #AIvideo #AIMusicVideo #AIFilmmaking

Kiber Alla

66,057 görüntüleme • 1 ay önce

xAI isn't playing around. They just released the Grok Imagine API, a unified video + image generation toolkit, and it's already sitting at #1 on the Artificial Analysis Video Arena for both Text-to-Video AND Image-to-Video. It's beating: ● Google's Veo 3.1 & Veo 3 ● OpenAI's Sora 2 ● Runway Gen-4.5 ● Kling 2.5 Turbo The Numbers Don't Lie: ● 64.1% win rate against Runway Aleph in blind human evaluations ● 57% win rate against Kling o1 ● Best-in-class latency. Sub-20 second generation for 720p, 8-second videos. (up to 15-second video) ● Native audio generation baked right into video output (dialogue, music, sound effects, all synced) What Makes It Different It's built for real creative workflows: ✅ Text-to-video AND image-to-video in one API ✅ Video editing with prompt-based controls (add/remove objects, restyle scenes) ✅ Camera controls: zoom, pan, timelapse, pull-back ✅ Style transfers: cyberpunk, watercolor, anime, you name it ✅ Performance animation: map your movements onto characters ✅ Native audio-video sync (no post-production needed) Why the focus on speed and cost? The partner feedback that shaped this: "Quality alone isn't enough if latency and cost make iteration painful." So xAI optimized for all three. Speed. Cost. Quality. Already Integrated With: ● fal. ai ● ComfyUI ● InVideo ● Flora ● HeyGen xAI went from underdog to chart-topper. The Grok Imagine API is fast, affordable, and genuinely production-ready. If you're building anything with AI video, this just became the one to beat.

tetsuo

18,325 görüntüleme • 7 ay önce

This week is already so hot. 🔥 Massive release from Decart : Lucy 2.0 a World Editing Model running at 1080p, 30FPS in realtime. This is truly exciting, the era of real-time generative reality is here. We are moving from watching AI video to living inside AI video. A breakthrough model capable of transforming the visual world in real-time. Moving beyond offline rendering, Lucy 2.0 delivers high-fidelity 1080p video generation with near-zero latency. Lucy 2.0 literally "redraws" the entire world pixel-by-pixel, while you are watching it. e.g. If you want to be an anime character, it doesn't just put a mask on you. It turns your skin into anime skin, your hair into anime hair, and the lighting in your room into anime lighting. Lucy 2.0 is also trained to stop the generated video from slowly falling apart over time, so the same stream can run much longer without faces and details drifting. So why is this a "Massive Deal"? Traditional AI video-generation model takes a prompt, you wait 10–20 minutes, and the computer "bakes" a video for you. You couldn't touch it or change it while it was happening. But Lucy 2.0 works like a mirror. It happens in real-time (30 frames per second). There is no waiting. You move your hand, the AI character moves its hand instantly. The craziest part isn't the visuals; it's the physics. Usually, AI hallucinations are glitchy—hands merge into faces, walls melt. Lucy 2.0 understands how the world works without being told. It knows that if you take off a helmet, there is hair underneath. It knows that if you splash water, droplets fly. It learned "physics" just by watching millions of videos. The physical behavior you see emerges from learned visual dynamics, not from engineered geometry or explicit physics engines. Their official technical report explicitly states that the model does not use traditional 3D engines, depth maps, or wireframes. It is a "pure diffusion model."

Rohan Paul

12,761 görüntüleme • 7 ay önce