Загрузка видео...

Не удалось загрузить видео

На главную

We are publishing our second deep dive today as a follow-up post on SLAM and VIO in egocentric tracking. We go deep into the sensor tradeoffs b/w global shutter and rolling shutter and their implications on SLAM / VIO - specifically how the way the camera reads each frame...

17,393 просмотров • 5 месяцев назад •via X (Twitter)

Комментарии: 11

Фото профиля FPV Labs
FPV Labs5 месяцев назад

Read the full essay here →

Фото профиля Shubhanshu Khatana
Shubhanshu Khatana5 месяцев назад

such a great read! the best part is its not so trivial yet one of the most important choice when it comes to choosing hardware for downstream SLAM/VIO pipelines

Фото профиля Daniel Lu
Daniel Lu5 месяцев назад

You can also do a combination of continuous time and discrete time approaches. It's often sufficient to use a first order approximation to "de-warp" the distorted features using velocity estimates from VIO or even just IMU. Later loop closures won't affect it too much.

Фото профиля FPV Labs
FPV Labs5 месяцев назад

This is very true. Instead of estimating a pose for every row of the image, you could also unwarp features in the image to compensate for the motion effect. This is what is typically done in consumer mobile phones these days, where the effects of rolling shutter may be compensated through some computational photography, so the image may look rolling-shutter-artifact-free in many cases.

Фото профиля Sam U
Sam U5 месяцев назад

quick question, with the rolling shutter example above, did we introduced a depth sensor to get the depth data or was it purely based off of the video ?

Фото профиля FPV Labs
FPV Labs5 месяцев назад

The test in the video is simply a feature tracking comparison in rolling shutter vs global shutter mono video. No state estimation or depth or any other data is involved.

Фото профиля Sam U
Sam U5 месяцев назад

👀👀

Фото профиля Bruno Santos🇵🇹
Bruno Santos🇵🇹5 месяцев назад

Is it Lie algebra ?

Фото профиля FPV Labs
FPV Labs5 месяцев назад

Lie algebra is a parameterization choice, particularly for non-linear optimization. The blog post itself just discusses simple SE(3) matrices in the setup, but the optimization of the underlying cost function for example - bundle adjustment - usually involves non-linear optimization in different parameter space - the most common choice is the lie algebra space

Фото профиля Annu Shekhawat
Annu Shekhawat5 месяцев назад

💪🔥🔥

Фото профиля Satpal Singh Rathore
Satpal Singh Rathore5 месяцев назад

Putting out learnings in open🚀

Похожие видео

Once we started to work with large global retailers, we needed a better way to scale this process. Ideally, the staff at the store could do this themselves — rather than us flying our team across the world — and then we could lower the cost and timelines. So we built a self-serve version of our survey app, with a tutorial mode designed for beginners. Over time, we collected millions of data points, and so we were able to develop an algorithm which would auto-correct mistakes. In other words, if the surveyor accidentally placed their ground-truth location in the wrong place on the map, we could use our algorithms to detect it, and correct it. So now we have WiFi, and with and our efforts on producing a high quality survey, we have the best WiFi positioning available. With WiFi on its own, it’s achieving 3 meter accuracy. This is a great foundation to build on. WiFi + Motion data To refine this down to 1-meter accuracy, we realised that we could combine WiFi with the same technology behind self-driving cars and robotics: a motion system called SLAM (Simultaneous Localization and Mapping). SLAM uses the accelerometer, gyroscope and camera system to understand precise device motion. Imagine a car driving through a tunnel, using the motion since its last GPS ping to keep location accurate until it comes out the other side. On a phone, this technology is very reliable, and measures device motion with high precision. But SLAM is measuring motion within its own coordinate space, it’s not aligned with the real world. SLAM tracks the user’s relative motion, like “moved forward 2 meters, then turned left”, but does “forward” mean “north”, or some other direction? It’s not calibrated, so it could mean any location, any direction. We can’t rely on the compass to help us out with this, because phone compasses are notoriously incorrect — everyone knows the frustration of being sent the wrong way down a street. So our job was to align this motion data with the triangulation data we were receiving from WiFi. We designed an algorithm that could simulate every possibility, filter the unlikely scenarios, and hone in your location, using WiFi as an anchor. So WiFi gives us the initial blue dot, SLAM gives us motion, and as the user starts walking and we receive more data, our algorithms can refine location accuracy down to a consistent 1-meter accuracy. We’ve tested these algorithms in many locations, on hundreds of hours of ground-truth data:

Andrew Hart

91,047 просмотров • 1 год назад