Video yükleniyor...
Video Yüklenemedi
Excited to introduce Intelligent Go-Explore: Foundation model (FM) agents have the potential to be invaluable, but struggle to learn hard-exploration tasks! Our new algorithm drastically improves their exploration abilities via Go-Explore + FM intelligence. Led by Cong Lu🧵1/
67,116 görüntüleme • 2 yıl önce •via X (Twitter)
10 Yorum

We revisit Go-Explore, a powerful exploration algorithm based on archiving discovered states & returning to and exploring from promising states. With the help of domain-specific heuristics, Go-Explore (Nature, 2021) achieved superhuman performance on Atari games and robotics! 2/

Key insight: leverage the ability of Foundation Models to judge interestingness in place of heuristics, extending Go-Explore to automatically tackle virtually any problem. IGE captures serendipity on the fly, including discoveries that are impossible to predict ahead of time! 3/

We evaluate IGE on a variety of text-based tasks that require search and exploration, with impressive results. For example, in Game of 24, a mathematical reasoning problem, IGE achieves 100% success rate 70.8% faster 📈 than classic graph search. 4/

On the hard TextWorld Coin Collector problem, a classic text-agent benchmark, IGE succeeds where prior SOTA FM agents completely fail. It also outperforms on medium and easier problems. 5/

IGE combines the tremendous strengths of foundation models with the powerful Go-Explore algorithm, opening up a new frontier of research into creating more generally capable agents with strong exploration capabilities. We're excited to see what new problems this could solve! 6/

IGE was driven by @cong_ml, in collaboration with @shengranhu and myself. 👩💻All code is open-sourced at: 🌐Website: 📄Paper: 7/7

@cong_ml Bigger environments? NetHack?

@cong_ml Coming soon to a theater near you!

@cong_ml Really cool work!

@cong_ml Highly informative



