Loading video...

Video Failed to Load

Go Home

Learning policies via imitation is extremely potent, but making sure those policies will generalize to out of distribution settings is still very hard. SAILOR proposes a solution in learning to search via a learned world model, which outperforms existing imitation approaches. Gokul Swamy, Arnav Jain , and Vibhakar Mohta...

21,507 views • 1 year ago •via X (Twitter)

7 Comments

Chris Paxton's profile picture
Chris Paxton1 year ago

For links and more + email updates you can check out the substack post for this episode:

Ipsita Praharaj's profile picture
Ipsita Praharaj2 months ago

Pretty cool!!

Saksham Jindal's profile picture
Saksham Jindal1 year ago

@arnavkj95 came a looooong way. Great work!

Chetan's profile picture
Chetan1 year ago

I really liked this paper!

VinayK's profile picture
VinayK2 months ago

Thanks for sharing

Yash Butala's profile picture
Yash Butala2 months ago

Wow!

Vaibhav Maheshwari's profile picture
Vaibhav Maheshwari2 months ago

Brilliant stuff

Related Videos