LeRobot's banner
LeRobot's profile picture

LeRobot

@LeRobotHF19,122 subscribers

~ Lowering the barrier to entry for robotics ~ Crafted with care by @HuggingFace 🤗 Join our discord: https://t.co/Sx2jdT0jeF

Shorts

We built a bipedal robot for about $2,500. A real, mostly 3D-printed robot you can build, repair, simulate, train, and control. Today we’re releasing LeRobot Humanoid: an open robot-learning platform with hardware, runtime, identification tools, and training environments. Blog post: Repo:

We built a bipedal robot for about $2,500. A real, mostly 3D-printed robot you can build, repair, simulate, train, and control. Today we’re releasing LeRobot Humanoid: an open robot-learning platform with hardware, runtime, identification tools, and training environments. Blog post: Repo:

200,072 Aufrufe

Hey Sunday, we also tried that! But ACT2 seems to work way better, congrats on the release!

Hey Sunday, we also tried that! But ACT2 seems to work way better, congrats on the release!

85,357 Aufrufe

LeRobot can now see in 3D 👀🤖 Depth cameras feed straight into training data end-to-end in v0.6.0: 12-bit precision preserved through encoding, or fully lossless raw storage if you'd rather skip compression altogether. Same release also opens up every video encoding parameter for RGB, plus a live benchmark leaderboard to help you tune it. This exact depth pipeline powers Stanford Vision and Learning Lab's BEHAVIOR-1K 2026 dataset: 20,000 simulated manipulation episodes, depth included. More on how it works below:

LeRobot can now see in 3D 👀🤖 Depth cameras feed straight into training data end-to-end in v0.6.0: 12-bit precision preserved through encoding, or fully lossless raw storage if you'd rather skip compression altogether. Same release also opens up every video encoding parameter for RGB, plus a live benchmark leaderboard to help you tune it. This exact depth pipeline powers Stanford Vision and Learning Lab's BEHAVIOR-1K 2026 dataset: 20,000 simulated manipulation episodes, depth included. More on how it works below:

28,841 Aufrufe

🤖 NVIDIA’s Gr00t N1.5 is now available in LeRobot! This is the result of a great collaboration between the Hugging Face LeRobot team and NVIDIA Robotics ! Gr00t N1.5 highlights: 🦾 Cross-embodiment foundation model for robots 🧠 Multimodal inputs: vision, language, and proprioception 🪛Tested on the Libero benchmark and real-world hardware tasks 🌍Trained on real robot, synthetic, and internet-scale video data ⚙️ Flow matching action transformer for action prediction

🤖 NVIDIA’s Gr00t N1.5 is now available in LeRobot! This is the result of a great collaboration between the Hugging Face LeRobot team and NVIDIA Robotics ! Gr00t N1.5 highlights: 🦾 Cross-embodiment foundation model for robots 🧠 Multimodal inputs: vision, language, and proprioception 🪛Tested on the Libero benchmark and real-world hardware tasks 🌍Trained on real robot, synthetic, and internet-scale video data ⚙️ Flow matching action transformer for action prediction

115,194 Aufrufe

🤖 Another zero-shot reward model is now in LeRobot: ROBOMETER. A general-purpose, zero-shot video-language reward model from University of South Carolina, UT Dallas, Massachusetts Institute of Technology (MIT), University of Washington, Ai2, and NVIDIA that predicts frame-level task progress. Trained on 1M+ trajectories from 21 robot embodiments, generalizes zero-shot to unseen tasks, scenes, and robots. 2.4–4.5x better downstream success rates across online RL, offline RL, data filtering, failure detection, and data retrieval for IL. Project: Paper:

🤖 Another zero-shot reward model is now in LeRobot: ROBOMETER. A general-purpose, zero-shot video-language reward model from University of South Carolina, UT Dallas, Massachusetts Institute of Technology (MIT), University of Washington, Ai2, and NVIDIA that predicts frame-level task progress. Trained on 1M+ trajectories from 21 robot embodiments, generalizes zero-shot to unseen tasks, scenes, and robots. 2.4–4.5x better downstream success rates across online RL, offline RL, data filtering, failure detection, and data retrieval for IL. Project: Paper:

32,625 Aufrufe

MolmoAct2 is landing in LeRobot! Ai2's open Action Reasoning Model combines a Molmo2-ER vision-language backbone with a flow-matching continuous action expert to predict robot action chunks from images, language instructions, and proprioceptive state. An open robot foundation model built for real-world control, with strong out-of-the-box performance and easy fine-tuning in LeRobot. Pick-and-place inference running on NVIDIA DGX Spark! Blog: Paper: Thanks to Ai2 Jiafei Duan Haoquan Fang

MolmoAct2 is landing in LeRobot! Ai2's open Action Reasoning Model combines a Molmo2-ER vision-language backbone with a flow-matching continuous action expert to predict robot action chunks from images, language instructions, and proprioceptive state. An open robot foundation model built for real-world control, with strong out-of-the-box performance and easy fine-tuning in LeRobot. Pick-and-place inference running on NVIDIA DGX Spark! Blog: Paper: Thanks to Ai2 Jiafei Duan Haoquan Fang

25,059 Aufrufe

With two SO100 leaders, and followers, you can now create a bimanual setup out of the box in LeRobot! 🦾 And we’re just getting started, more robots are on the way! 🤖 Which one would you like to see next? ⏩

With two SO100 leaders, and followers, you can now create a bimanual setup out of the box in LeRobot! 🦾 And we’re just getting started, more robots are on the way! 🤖 Which one would you like to see next? ⏩

17,805 Aufrufe

Videos