Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Introducing Axis Dataset V1 - the simulation dataset for scalable robot manipulation. Can noisy, crowdsourced simulation data support embodied pretraining? Yes—with enough coverage and diversity. Continual pretraining on AXIS dataset V1 lifts π0.5 from 83.9% to 88.8% on LIBERO-Plus, with performance improving consistently as the pretraining data scales from...

96,538 Aufrufe • vor 1 Monat •via X (Twitter)

34 Kommentare

Profilbild von Stone Tao
Stone Taovor 1 Monat

don’t want to put too much shade on the progress (i think sim datasets can be good ideas if executed well!), but improving on libero (which is also a sim benchmark) means very little unfortunately

Profilbild von Thor
Thorvor 1 Monat

Let's goo 🔥

Profilbild von Hazzy
Hazzyvor 1 Monat

you edit too much intern

Profilbild von KABILA
KABILAvor 1 Monat

Cookingg

Profilbild von Jontu
Jontuvor 1 Monat

This is a strong proof that diverse simulation data can meaningfully improve embodied AI.

Profilbild von boyBrands_
boyBrands_vor 1 Monat

He has arrived🔥

Profilbild von Sami
Samivor 1 Monat

So much excited for this

Profilbild von Apurbo
Apurbovor 1 Monat

We are so excited about this 🤘

Profilbild von AL Shifat
AL Shifatvor 1 Monat

Axis going to make history

Profilbild von dviet.eth
dviet.ethvor 1 Monat

This is exactly the kind of infrastructure Physical AI needs. Quality data compounds, and so does robot intelligence. Excited to see Axis Dataset V1 keep growing 😆

Profilbild von Noodle
Noodlevor 1 Monat

cool to see those teleop sessions actually pushing the libero-plus numbers up

Profilbild von Acee (✱,✱)
Acee (✱,✱)vor 1 Monat

Keep cooking 👌

Profilbild von Piggie 🧡|Bird
Piggie 🧡|Birdvor 1 Monat

Overall success rate increased by 5.8%.

Profilbild von 0xmowra.eth
0xmowra.ethvor 1 Monat

Axis V1 dataset boosting robot learning with crowdsourced simulation data impressive

Profilbild von Red
Redvor 1 Monat

Great work

Profilbild von | MinhDucc | Shark
| MinhDucc | Sharkvor 1 Monat

Axis LFG 😍

Profilbild von sasafonda✱,✱)
sasafonda✱,✱)vor 1 Monat

gaxis

Profilbild von namnew
namnewvor 1 Monat

gAxis

Profilbild von VINAY (✱,✱)
VINAY (✱,✱)vor 1 Monat

Axis cooking 🔥

Profilbild von savagezz
savagezzvor 1 Monat

cooking something

Profilbild von Hoangdong (✱,✱)
Hoangdong (✱,✱)vor 1 Monat

gAxis

Profilbild von clover
clovervor 1 Monat

move gaxis

Profilbild von vinhle (✱,✱) 🍚 ⛓ ./
vinhle (✱,✱) 🍚 ⛓ ./vor 1 Monat

Exciting to see the potential of crowdsourced simulation data in robotics, and impressive results from continual pretraining on Axis Dataset V1.

Profilbild von Aurora Lux
Aurora Luxvor 1 Monat

AXIS Dataset V1 is a strong step forward for embodied AI—scalable, diverse simulation data is proving to be a powerful foundation for better robot learning. 🤖🚀

Profilbild von cbeta.base.eth
cbeta.base.ethvor 1 Monat

lets goo team fight fight fight

Profilbild von rara
raravor 1 Monat

gaxis lfg

Profilbild von Linh (✱,✱)
Linh (✱,✱)vor 1 Monat

Axis Dataset V1 is impressive

Profilbild von Nova
Novavor 1 Monat

The bot is terrible. Please update it. Its performance is like a primitive human living in a cave 😤

Profilbild von MAS_Enpici (✱,✱)
MAS_Enpici (✱,✱)vor 1 Monat

big update and new insight

Profilbild von yellowopus
yellowopusvor 1 Monat

lets cooking bro

Profilbild von KADAFER
KADAFERvor 1 Monat

let him cook 🔥

Profilbild von dondonat (✱,✱)
dondonat (✱,✱)vor 1 Monat

gaxis keep building

Profilbild von sagar
sagarvor 1 Monat

Let's goo 🙌

Profilbild von Gujil Ruipa
Gujil Ruipavor 1 Monat

Open datasets like this can accelerate robotics research through collaborative innovation

Ähnliche Videos

In my past research experience, finding or developing an appropriate simulation environment, dataset, and benchmark has always been a challenge. Missing features, limited support, or unexpected bugs often occupied my days and nights. Moreover, current simulation platforms are relatively fragmented—making it challenging to replicate the success of the RT-X dataset in unifying community efforts. Introducing RoboVerse, we provide a unified platform, dataset, and benchmark for scalable and generalizable robot learning. We hope to build a shared foundation to combine the community efforts. RoboVerse includes: MetaSim: We carefully designed a configuration system and a universal interface to align current robotic simulators. With MetaSim, you can use any simulator with the same code—bringing together the community’s diverse efforts under one framework! RoboVerse Dataset and Benchmark: We unify popular simulation environments and benchmarks into a single cohesive system and introduce the RoboVerse dataset—a large-scale, high-quality synthetic dataset. Additionally, we propose a standardized benchmark across both imitation learning and reinforcement learning. A cool feature enabled by our unified framework: Hybrid Simulation! You can now integrate physics engines and renderers from different simulators—e.g., using MuJoCo precise physics with Isaac photorealistic rendering. This not only elevates simulation fidelity but also significantly enhances real-world transfer performance across complex robotic applications. Hopefully, our team’s efforts could serve the robotic community to thrive vibrantly in the years to come. RoboVerse is open-sourced🥳!!! Project Page: Documentation: Github Repo: Paper:

Haoran Geng

84,318 Aufrufe • vor 1 Jahr

I don’t know if we live in a Matrix, but I know for sure that robots will spend most of their lives in simulation. Let machines train machines. I’m excited to introduce DexMimicGen, a massive-scale synthetic data generator that enables a humanoid robot to learn complex skills from only a handful of human demonstrations. Yes, as few as 5! DexMimicGen addresses the biggest pain point in robotics: where do we get data? Unlike with LLMs, where vast amounts of texts are readily available, you cannot simply download motor control signals from the internet. So researchers teleoperate the robots to collect motion data via XR headsets. They have to repeat the same skill over and over and over again, because neural nets are data hungry. This is a very slow and uncomfortable process. At NVIDIA, we believe the majority of high-quality tokens for robot foundation models will come from simulation. What DexMimicGen does is to trade GPU compute time for human time. It takes one motion trajectory from human, and multiplies into 1000s of new trajectories. A robot brain trained on this augmented dataset will generalize far better in the real world. Think of DexMimicGen as a learning signal amplifier. It maps a small dataset to a large (de facto infinite) dataset, using physics simulation in the loop. In this way, we free humans from babysitting the bots all day. The future of robot data is generative. The future of the entire robot learning pipeline will also be generative. 🧵

Jim Fan

165,246 Aufrufe • vor 1 Jahr

🔥 JUST IN: Open-source robotics dataset from 100% real-world scenarios! 🤯 Chinese robotics company AGIBOT just released AGIBOT WORLD 2026, an open-source dataset systematically covering key embodied AI research directions. Built entirely from real-world environments: commercial spaces, and homes. Collected using AGIBOT G2 robots in free-form collection mode, providing structured, accurately annotated, high-quality data. Digital twin technology creates 1:1 scale replicas in simulation matching the real environments. Both real-world and simulation data are open-sourced. The AGIBOT G2 platform collects multiple data types simultaneously: RGB(D) cameras, tactile sensors, force sensors, LiDAR, IMU, and full-body joint states. Whole-body control coordinates arms, waist, and hands for complex tasks. First-person teleoperation lets operators control the robot from its perspective. The tasks covered are fine-grained manipulation, ultra-long-horizon tasks, spatial navigation, dual-arm coordination, and multi-agent/human-robot collaboration. The dataset includes error-recovery trajectories with annotations. Most datasets only show successful demonstrations. AGIBOT includes failures and how the robot recovers, teaching models how to handle mistakes. After collection, data is tested through policy training and real-robot deployment to ensure quality. Then processed through industrial quality control with multiple screening and cleaning rounds. Making it open-source accelerates embodied AI research by giving researchers access to high-quality real-world robot data at scale. 🇨🇳 Learn more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

40,583 Aufrufe • vor 5 Monaten

Synthetic data will provide the next trillion tokens to fuel our hungry models. I'm excited to announce MimicGen: massively scaling up data pipeline for robot learning! We multiply high-quality human data in simulation with digital twins. Using 50,000 training episodes across 18 tasks, multiple simulators, and even in the real-world! The idea is simple: 1. Humans tele-operate the robot to complete a task. It is extremely high-quality but also very slow and expensive. 2. We create a digital twin of the robot and the scene in high-fidelity, GPU-accelerated simulation. 3. We can now move objects around, replace with new assets, and even change the robot hand - basically augment the training data with procedural generation. 4. Export the successful episodes, and feed that to a neural network! You now have an near-infinite stream of data. One of the key reasons that robotics lags far behind other AI fields is the lack of data: you cannot scrape control signals from the internet. They simply don't exist in-the-wild. MimicGen shows the power of synthetic data and simulation to keep our scaling laws alive. I believe this principle apply beyond robotics. We are quickly exhausting the high-quality, real tokens from the web. Artificial intelligence from artificial data will be the way forward. We are big fans of the OSS community. As usual, we open-source everything, including the generated dataset! - Website: - Paper: - Dataset is hosted on HuggingFace (thanks AK!!): - Code: MimicGen is led by Ajay Mandlekar, deep dive in the thread:

Jim Fan

332,238 Aufrufe • vor 2 Jahren

Excited to announce GR00T N1, the world’s first open foundation model for humanoid robots! We are on a mission to democratize Physical AI. The power of general robot brain, in the palm of your hand - with only 2B parameters, N1 learns from the most diverse physical action dataset ever compiled and punches above its weight: - Real humanoid teleoperation data. - Large-scale simulation data: we are open-sourcing 300K+ trajectories! - Neural trajectories: we apply SOTA video generation models to “hallucinate” new synthetic data that features accurate physics in pixels. Using Jensen’s words, “systematically infinite data”! - Latent actions: we develop novel algorithms to extract action tokens from in-the-wild human videos and neural generated videos. GR00T N1 is a single end-to-end neural net, from photons to actions: - Vision-Language Model (System 2) that interprets the physical world through vision and language instructions, enabling robots to reason about their environment and instructions, and plan the right actions. - Diffusion Transformer (System 1) that “renders” smooth and precise motor actions at 120 Hz, executing the latent plan made by System 2. We deploy N1 on GR1 robot, 1X Neo robot, and a large collection of simulation benchmarks. N1 achieves up to +30% boost in diverse manipulation tasks for household and industrial settings. While humanoid robots are the main focus of N1, our model also supports cross-embodiment. We finetune it to work on the $110 HuggingFace LeRobot SO100 robot arm! Open robot brain runs on open hardware. Sounds just right. Let’s solve robotics, together, one token at a time. Links to our Whitepaper, Github repo, HuggingFace model, and open dataset page in the thread: 🧵

Jim Fan

467,416 Aufrufe • vor 1 Jahr

*New Paper on AI & Democracy* Imagine two approaches to democracy. The one we have today, where citizens choose a professional politician to represent them and others. Or an augmented form of democracy, where each citizen controls a personalized AI that helps them participate in thousands of nuanced decisions. This second approach is the idea of Augmented Democracy I introduced six years ago at TED. In our latest paper we explore a simplified version of Augmented Democracy by combining off-the-shelf LLMs, such as ChatGPT, with data collected using a collaborative government program builder. This was an online game where people build a personalized government program using proposals extracted from the programs of the candidates of the 2022 presidential election in Brazil. So how accurate are these augmented forms of democracy? Imagine a user who gave us 40 answers. We can use the first 20 to fine-tune a model that we can test using the 20 answers the model didn’t see. We can then compare the accuracy of these predictions with the ones obtained by a “bundle” rule, which assumes that users that self-reported to be from the left or right always chose the proposals from the candidate that shares their political identity. This showed us that LLMs were more accurate at predicting policy preferences than the bundle rule, meaning that the preferences captured in the participation data were more nuanced than a left-right axis, and that the LLMs can capture some of that nuance. Also, the LLMs can choose among policies coming from the same candidate, which is something that we cannot do using a bundle rule. But can these LLMs help us complete the aggregate preferences of the population? Direct or unbundled forms of participation can result in incomplete data when people answer only a fraction of all questions. In our paper, we simulate this incompleteness by sampling the full dataset. We ask how close we can get to the full dataset by using a random sample, or a random sample augmented by predictions made by these LLMs. Overall, we find that LLM-augmented data gets much closer to the full dataset than a pure random sample. These results do not mean that augmented democracy technology is ready, but they means we are in a much better place to continue exploring this idea than six years ago. This paper was a collaborative effort with Jairo Gudino, PhD student at CCL at the University of Toulouse Capitole and Umberto Grandi from IRIT also at the University of Toulouse Capitole. We hope you find these results insightful!

César A. Hidalgo

26,915 Aufrufe • vor 1 Jahr

Exciting updates on Project GR00T! We discover a systematic way to scale up robot data, tackling the most painful pain point in robotics. The idea is simple: human collects demonstration on a real robot, and we multiply that data 1000x or more in simulation. Let’s break it down: 1. We use Apple Vision Pro (yes!!) to give the human operator first person control of the humanoid. Vision Pro parses human hand pose and retargets the motion to the robot hand, all in real time. From the human’s point of view, they are immersed in another body like the Avatar. Teleoperation is slow and time-consuming, but we can afford to collect a small amount of data. 2. We use RoboCasa, a generative simulation framework, to multiply the demonstration data by varying the visual appearance and layout of the environment. In Jensen’s keynote video below, the humanoid is now placing the cup in hundreds of kitchens with a huge diversity of textures, furniture, and object placement. We only have 1 physical kitchen at the GEAR Lab in NVIDIA HQ, but we can conjure up infinite ones in simulation. 3. Finally, we apply MimicGen, a technique to multiply the above data even more by varying the *motion* of the robot. MimicGen generates vast number of new action trajectories based on the original human data, and filters out failed ones (e.g. those that drop the cup) to form a much larger dataset. To sum up, given 1 human trajectory with Vision Pro -> RoboCasa produces N (varying visuals) -> MimicGen further augments to NxM (varying motions). This is the way to trade compute for expensive human data by GPU-accelerated simulation. A while ago, I mentioned that teleoperation is fundamentally not scalable, because we are always limited by 24 hrs/robot/day in the world of atoms. Our new GR00T synthetic data pipeline breaks this barrier in the world of bits. Scaling has been so much fun for LLMs, and it's finally our turn to have fun in robotics! We are building tools to enable everyone in the ecosystem to scale up with us. Links in thread:

Jim Fan

364,670 Aufrufe • vor 2 Jahren