正在加载视频...
视频加载失败
How can we help *any* image-input policy generalize better? 👉 Meet PEEK 🤖 — a framework that uses VLMs to decide *where* to look and *what* to do, so downstream policies — from ACT, 3D-DA, or even π₀ — generalize more effectively! 🧵
13,295 次观看 • 8 个月前 •via X (Twitter)
0 条评论
暂无评论
原始帖子的评论将显示在这里
