正在加载视频...
视频加载失败
Everyone in Embodied AI is talking about Vision-Language-Action (VLA) models. Almost no one is talking about the physical nightmare of collecting the data to train them. You can't scrape a kitchen table or a warehouse shelf from a web browser. To get to millions of hours of diverse, real-world... show more
26,385 次观看 • 2 个月前 •via X (Twitter)
0 条评论
暂无评论
原始帖子的评论将显示在这里
