正在加载视频...
视频加载失败
3D editing is hard: you need to ground an image + instruction and generate a faithful 3D shape in one forward pass -- no test-time optimization. So, we steer pretrained image-to-3D representations to do text-guided 3D edits; no massive 3D edit-pair dataset needed. Key trap: the “no-edit” solution is... show more
0 条评论
暂无评论
原始帖子的评论将显示在这里
