正在加载视频...
视频加载失败
VLA policies learn generalist robot behaviors from massive teleoperation datasets, hoping that the right behavior emerges. But they rarely use perception during training or inference: powerful foundation models of 3D geometry, semantics, or human motion are ignored. TimSong, Long Le, and our GRASP Laboratory team introduce Omniguide, based on... show more
47,950 次观看 • 5 个月前 •via X (Twitter)
0 条评论
暂无评论
原始帖子的评论将显示在这里
