Loading video...

Video Failed to Load

Go Home

Human Archive builds datasets to model sensorimotor intelligence for robotics and world models. HA-Multi is the largest dataset of manual labor tasks aligning egocentric RGB, stereo depth (active IR), tactile gloves, body IMUs, and wrist cameras. Congrats on the launch, Raj Patel, Shloke Patel, Samay Maini, and Rushil Agarwal!

110,791 views • 6 months ago •via X (Twitter)

34 Comments

Runtime's profile picture
Runtime6 months ago

Onwards and upwards!

Panda Hodler (YC Intern)'s profile picture
Panda Hodler (YC Intern)6 months ago

This seems to equip robots with "human-level tactile and visual nerves." The team's approach is very interesting i think😆

WildPinesAI's profile picture
WildPinesAI6 months ago

tactile force maps aligned with vision is the missing dataset every robotics foundation model is starving for

Samay Maini's profile picture
Samay Maini6 months ago

let’s archive the world for embodied intelligence!

Kai's profile picture
Kai6 months ago

This stack is wild, egocentric video plus tactile plus IMU in one dataset is exactly what makes sim to real less brittle. Fine tuning policies from this should beat pure vision only sets.

Maahir Sachdev's profile picture
Maahir Sachdev6 months ago

congrats champ

Kyriakos's profile picture
Kyriakos6 months ago

AI needs real world datasets

Manideep's profile picture
Manideep6 months ago

this is goated startup in this batch

Jai Bhatia's profile picture
Jai Bhatia6 months ago

Congrats boys!

Skyler Chan's profile picture
Skyler Chan6 months ago

congrats raj!

BTCaveman.'s profile picture
BTCaveman.6 months ago

Lfg boys! Kill it ! Proud to see my Indians together like this

The Humanoid Hub's profile picture
The Humanoid Hub6 months ago

Models are hungry!!

Mohd Tanveer's profile picture
Mohd Tanveer6 months ago

Exciting step forward for embodied AI and robotics. Congrats to the team! 🚀

Chain Alpha's profile picture
Chain Alpha6 months ago

Congrats!

Rushil Agarwal's profile picture
Rushil Agarwal6 months ago

Super excited

sreekar nagulapalli's profile picture
sreekar nagulapalli6 months ago

let’s goooo

Rithvik's profile picture
Rithvik6 months ago

Goats

Sanjay Adhikesaven's profile picture
Sanjay Adhikesaven6 months ago

congrats!

Abhishek Vadapalli's profile picture
Abhishek Vadapalli6 months ago

Congrats @babugi28!!

parag parikh's profile picture
parag parikh6 months ago

Congratulations and Goodluck ...

Chawit's profile picture
Chawit6 months ago

looks amazing!

vishnu manoj's profile picture
vishnu manoj6 months ago

lfgg🚀

Testimony | Product Designer's profile picture
Testimony | Product Designer6 months ago

put enough sensors on a human doing manual labor and suddenly you have the most expensive dishwashing video ever recorded. congrats on the launch

justin wang's profile picture
justin wang6 months ago

congrats on the launch raj!!

Shamim Hossain's profile picture
Shamim Hossain6 months ago

Congratulations on the launch!

interstellar's profile picture
interstellar6 months ago

Great !

Mohamed Anis's profile picture
Mohamed Anis6 months ago

Excited to see how this enables the next generation of foundation models for physical AI.

The Daily_Ai's profile picture
The Daily_Ai6 months ago

Exciting stuff! The integration of all those data types will really push robotics forward. Can't wait to see what comes next!

Rayan's profile picture
Rayan6 months ago

Congrats !

Ali Taha Brown's profile picture
Ali Taha Brown6 months ago

Congrats team

Neel Sharma's profile picture
Neel Sharma6 months ago

BANG

Keshav Agarwal's profile picture
Keshav Agarwal6 months ago

We have build some apps that we would soon be launching on X and would YC to have a look if they find interesting

Veer Sahasi's profile picture
Veer Sahasi6 months ago

Awesome!!

Shrinandan Narayanan's profile picture
Shrinandan Narayanan6 months ago

Congrats guys!

Related Videos

🔥 JUST IN: Open-source robotics dataset from 100% real-world scenarios! 🤯 Chinese robotics company AGIBOT just released AGIBOT WORLD 2026, an open-source dataset systematically covering key embodied AI research directions. Built entirely from real-world environments: commercial spaces, and homes. Collected using AGIBOT G2 robots in free-form collection mode, providing structured, accurately annotated, high-quality data. Digital twin technology creates 1:1 scale replicas in simulation matching the real environments. Both real-world and simulation data are open-sourced. The AGIBOT G2 platform collects multiple data types simultaneously: RGB(D) cameras, tactile sensors, force sensors, LiDAR, IMU, and full-body joint states. Whole-body control coordinates arms, waist, and hands for complex tasks. First-person teleoperation lets operators control the robot from its perspective. The tasks covered are fine-grained manipulation, ultra-long-horizon tasks, spatial navigation, dual-arm coordination, and multi-agent/human-robot collaboration. The dataset includes error-recovery trajectories with annotations. Most datasets only show successful demonstrations. AGIBOT includes failures and how the robot recovers, teaching models how to handle mistakes. After collection, data is tested through policy training and real-robot deployment to ensure quality. Then processed through industrial quality control with multiple screening and cleaning rounds. Making it open-source accelerates embodied AI research by giving researchers access to high-quality real-world robot data at scale. 🇨🇳 Learn more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

40,583 views • 5 months ago

🚨 BREAKING: Microsoft's first robotics foundation model! 🤯 Microsoft just announced Rho-alpha (ρα), their first robotics model derived from the Phi series of vision-language models. Rho-alpha translates natural language commands into control signals for robotic systems performing bimanual manipulation tasks. Commands like "push the green button with the right gripper," "pull out the red wire," "flip the top switch on," or "turn the knob to position 5" get executed directly by dual-arm robots. What makes this different from standard vision-language-action (VLA) models is the additional modalities. Rho-alpha is a VLA+ model that adds tactile sensing to the perceptual mix, with plans to incorporate force feedback. On the learning side, the model is designed to continually improve during deployment by learning from human feedback. The training approach combines trajectories from physical demonstrations and simulated tasks with web-scale visual question answering data. Since teleoperation data is scarce and expensive, Microsoft is using NVIDIA Isaac Sim on Azure to generate physically accurate synthetic datasets via reinforcement learning. These simulated trajectories get combined with commercial and open physical demonstration datasets. The model is currently under evaluation on dual-arm setups and humanoid robots. Microsoft is opening an Early Access Program for organizations interested in evaluating Rho-alpha. Robots that can adapt to dynamic situations and human preferences are more useful in real environments and more trusted by the people operating them. Read more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

60,985 views • 7 months ago

JUST IN: Dyna Robotics just published one of the most important research papers in robotics this year. It could fundamentally change how robot foundation models are trained. A scaling law that transfers from human video to robot performance. Dyna-2 is out and it's 🔥 Here's what that means in plain terms. Dyna-2 was pre-trained on ONE MILLION hours of egocentric human video, 170 years of continuous human experience, cooking, folding, assembling, cleaning. And as that human data scaled, robot performance improved. Predictably. Monotonically. Across 39 tasks on two different robot embodiments the model had never seen. → 1,000 hours pre-training → 20% normalised task performance → 10,000 hours → 28% → 100,000 hours → 45% → 1,000,000 hours → 53% Human video exists at effectively unlimited scale. Every cook, every factory worker, every craftsperson wearing a camera is generating training data for future robots. But the finding that stunned even the researchers, world modeling is what makes the transfer work. A model trained to predict future video AND actions massively outperforms one trained on actions alone. Video is the new scaling axis for robotics. One more jaw-dropping data point. 13 minutes of teleoperation data was enough to fine-tune Dyna-2 to open a bottle cap using two five-fingered robot hands. The robots are coming, and they're learning from us directly :D Read more here: Congrats Jason Ma and team! ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

23,576 views • 1 month ago