Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Presenting DemoDiffusion: An extremely simple approach enabling a pre-trained 'generalist' diffusion policy to follow a human-demonstration for a novel task during inference One-shot human imitation *without* requiring any paired human-robot data or online RL 🙂 1/n

32,919 görüntüleme • 1 yıl önce •via X (Twitter)

8 Yorum

Homanga Bharadhwaj profil fotoğrafı
Homanga Bharadhwaj1 yıl önce

The key insight of DemoDiffusion is to start the denoising process for the diffusion policy with the re-targeted human hand trajectory (instead of starting from pure noise) This simple approach doesn't require fine-tuning/updating the diffusion policy in any way! 2/n

Homanga Bharadhwaj profil fotoğrafı
Homanga Bharadhwaj1 yıl önce

Results show that DemoDiffusion can perform tasks that the pre-trained diffusion policy (pi-0) fails at zero-shot, just from one human demonstration of the task! 3/n

Homanga Bharadhwaj profil fotoğrafı
Homanga Bharadhwaj1 yıl önce

We even see zero-shot generalization to objects different from what the human demonstration was shown on! This suggests DemoDiffusion is able to exploit the semantic/spatial generalization of the pre-trained diffusion policy - while guiding it based on the human demo 4/n

Homanga Bharadhwaj profil fotoğrafı
Homanga Bharadhwaj1 yıl önce

DemoDiffusion is made possible by @sungj1026 's amazing lead, and @shubhtuls 's precise insights on diffusion models @CMU_Robotics Code, Videos, Paper: (finally, thanks to @physical_int for pi0 and @geopavlakos @JitendraMalikCV et al. for HaMeR) n/n

Homanga Bharadhwaj profil fotoğrafı
Homanga Bharadhwaj1 yıl önce

@shubhtuls @CMU_Robotics @physical_int @geopavlakos @JitendraMalikCV Also check out this alternate thread from @sungj1026 on DemoDiffusion (n+1)/n

Ted Xiao profil fotoğrafı
Ted Xiao1 yıl önce

Nice work! Warm-starting the denoising progress with a human prior is very smart.

Himanshu Kumar profil fotoğrafı
Himanshu Kumar1 yıl önce

Perhaps true mastery lies in effortless adaptation, not rigid programming.

Arsen Ibragimov profil fotoğrafı
Arsen Ibragimov1 yıl önce

Thats clever, skipping the fine-tuning part is a flex

Benzer Videolar