Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Learning policies via imitation is extremely potent, but making sure those policies will generalize to out of distribution settings is still very hard. SAILOR proposes a solution in learning to search via a learned world model, which outperforms existing imitation approaches. Gokul Swamy, Arnav Jain , and Vibhakar Mohta...

21,507 Aufrufe • vor 1 Jahr •via X (Twitter)

7 Kommentare

Profilbild von Chris Paxton
Chris Paxtonvor 1 Jahr

For links and more + email updates you can check out the substack post for this episode:

Profilbild von Ipsita Praharaj
Ipsita Praharajvor 2 Monaten

Pretty cool!!

Profilbild von Saksham Jindal
Saksham Jindalvor 1 Jahr

@arnavkj95 came a looooong way. Great work!

Profilbild von Chetan
Chetanvor 1 Jahr

I really liked this paper!

Profilbild von VinayK
VinayKvor 2 Monaten

Thanks for sharing

Profilbild von Yash Butala
Yash Butalavor 2 Monaten

Wow!

Profilbild von Vaibhav Maheshwari
Vaibhav Maheshwarivor 2 Monaten

Brilliant stuff

Ähnliche Videos