Loading video...
Video Failed to Load
ReViV reconstructs viewer-centric human motion (body, hand, and gaze) and view-centric scene geometry (camera and depth) from a single egocentric RGB video in a unified feed-forward model. It formulates the task as learning the full joint probability distribution over multimodal signals, including RGB video, camera trajectory, gaze direction, full-body... show more
18,590 views • 2 months ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
