Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

📢 New Paper PointSt3R: Point Tracking through 3D Grounded Correspondence Can point tracking be re-formulated as pairwise frame correspondence solely? We fine-tuning MASt3R with dynamic correspondences and a visibility loss and achieve competitive point tracking results 1/3

10,498 Aufrufe • vor 10 Monaten •via X (Twitter)

4 Kommentare

Profilbild von Dima Damen ✈️ #ECCV2026
Dima Damen ✈️ #ECCV2026vor 10 Monaten

Through balancing static and dynamic correspondences, the model can maintain MASt3R's power in tracking static parts of the scene but also track dynamic points successfully. We use *no* temporal knowledge - only pairwise matching!

Profilbild von Dima Damen ✈️ #ECCV2026
Dima Damen ✈️ #ECCV2026vor 10 Monaten

Trained visibility head shows impressive performance in identifying camera motion and tracking dynamic objects jointly... Code and Models are out Work led by Rhodri Guerrier w @AdamWHarley 3/3

Profilbild von ζ Pedram ζ
ζ Pedram ζvor 10 Monaten

Ooh, kind of a new pixel flow

Profilbild von Himanshu Kumar
Himanshu Kumarvor 10 Monaten

Ah, Dima, that's a clever idea! Re-formulating point tracking is quite innovative; I'm eager to see how it plays out.

Ähnliche Videos

Wow. Recreating the Shawshank Redemption prison in 3D from a single video, in real time (!) Just read the MASt3R-SLAM paper and it's pretty neat. These folks basically built a real-time dense SLAM system on top of MASt3R, which is a transformer-based neural network that can do 3d reconstruction and localization from uncalibrated image pairs. The cool part is they don't need a fixed camera model -- it just works with arbitrary cameras -- think different focal lengths, sensor sizes, even handling zooming in video (FMV drone video anyone?!). If you've done photogrammetry or played with NeRFs you know that is a HUGE deal. They've solved some tricky problems like efficient point matching and tracking, plus they've figured out how to fuse point clouds and handle loop closures in real-time. Their system runs at about 15 FPS on a 4090 and produces both camera poses and dense geometry. When they know the camera calibration, they get SOTA results across several benchmarks, but even without calibration, they still perform well. What's interesting is the approach -- most recent SLAM work has built on DROID-SLAM's architecture, but these folks went a different direction by leveraging a strong 3D reconstruction prior. Seems to give them more coherent geometry, which makes sense since that's what MASt3R was designed for. For anyone who cares about monocular SLAM and 3D reconstruction, this feels like a significant step toward plug-and-play dense SLAM without calibration headaches -- perfect for drones, robots, AR/VR -- the works!

Bilawal Sidhu

704,318 Aufrufe • vor 1 Jahr