Video wird geladen...
Video konnte nicht geladen werden
Learning policies via imitation is extremely potent, but making sure those policies will generalize to out of distribution settings is still very hard. SAILOR proposes a solution in learning to search via a learned world model, which outperforms existing imitation approaches. Gokul Swamy, Arnav Jain , and Vibhakar Mohta... show more
21,507 Aufrufe • vor 1 Jahr •via X (Twitter)
7 Kommentare

Chris Paxtonvor 1 Jahr
For links and more + email updates you can check out the substack post for this episode:

Ipsita Praharajvor 2 Monaten
Pretty cool!!

Saksham Jindalvor 1 Jahr
@arnavkj95 came a looooong way. Great work!

Chetanvor 1 Jahr
I really liked this paper!

VinayKvor 2 Monaten
Thanks for sharing

Yash Butalavor 2 Monaten
Wow!

Vaibhav Maheshwarivor 2 Monaten
Brilliant stuff
