Загрузка видео...

Не удалось загрузить видео

На главную

🤔Applying Depth Estimation models directly to videos can result in inconsistency between frames. 💪Well, not anymore. 🔥ChronoDepth is a new approach to video depth estimation that focuses on achieving both accuracy within each frame and consistency across frames. 📜More👇

19,992 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 0

Нет доступных комментариев

Здесь появятся комментарии из оригинального поста

Похожие видео

Depth Any Video with Scalable Synthetic Data AI physicists and chemists continue to make strides in depth estimation from video. Check out this new paper featuring some impressive examples. See the thread for more details (unfortunately no code yet). Abstract: Video depth estimation has long been hindered by the scarcity of consistent and scalable ground truth data, leading to inconsistent and unreliable results. In this paper, we introduce Depth Any Video, a model that tackles the challenge through two key innovations. First, we develop a scalable synthetic data pipeline, capturing real-time video depth data from diverse game environments, yielding 40,000 video clips of 5-second duration, each with precise depth annotations. Second, we leverage the powerful priors of generative video diffusion models to handle real-world videos effectively, integrating advanced techniques such as rotary position encoding and flow matching to further enhance flexibility and efficiency. Unlike previous models, which are limited to fixed-length video sequences, our approach introduces a novel mixed-duration training strategy that handles videos of varying lengths and performs robustly across different frame rates 0 - even on single frames. At inference, we propose a depth interpolation method that enables our model to infer high-resolution video depth across sequences of up to 150 frames. Our model outperforms all previous generative depth models in terms of spatial accuracy and temporal consistency.

MrNeRF

27,428 просмотров • 1 год назад

Contact sheet prompting is the hottest AI video technique right now 🤯 One image in → 6 consistent frames out → cinematic video ads in minutes. But everyone's doing it manually. I automated the entire workflow in n8n + Airtable. Here's why contact sheet prompting is blowing up: You give AI one reference image, and it generates a grid of consistent shots — same face, same outfit, different angles. Instant storyboarding, full creative control, no photoshoots. The problem? It's super tedious: → Write the prompt manually → Generate the contact sheet → Crop each frame by hand → Feed frames into a video model one by one → Repeat for every product This n8n automation handles all of it: → Upload character image + product image → AI analyzes both and writes the contact sheet prompt → Nano Banana Pro generates a 6-frame grid → System extracts each frame automatically → Kling 2.5 generates smooth transitions between frames → You get 5 video clips ready to stitch Approval checkpoints at every stage, no surprises. What lands in your Airtable: → AI-generated creative prompt → Core hero image (model + product) → 6-frame contact sheet → 5 cinematic video clips → Full control before each generation step Contact sheet prompting on autopilot. I filmed a 20 minute Loom video showing you exactly how I set it up. Want the Loom + the complete n8n workflow + Airtable base? > Comment "SHEET" > Like this post And I'll send it over (must be following so I can DM)

Mike Futia

53,686 просмотров • 8 месяцев назад