Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Human Archive builds datasets to model sensorimotor intelligence for robotics and world models. HA-Multi is the largest dataset of manual labor tasks aligning egocentric RGB, stereo depth (active IR), tactile gloves, body IMUs, and wrist cameras. Congrats on the launch, Raj Patel, Shloke Patel, Samay Maini, and Rushil Agarwal!

110,791 Aufrufe • vor 6 Monaten •via X (Twitter)

34 Kommentare

Profilbild von Runtime
Runtimevor 6 Monaten

Onwards and upwards!

Profilbild von Panda Hodler (YC Intern)
Panda Hodler (YC Intern)vor 6 Monaten

This seems to equip robots with "human-level tactile and visual nerves." The team's approach is very interesting i think😆

Profilbild von WildPinesAI
WildPinesAIvor 6 Monaten

tactile force maps aligned with vision is the missing dataset every robotics foundation model is starving for

Profilbild von Samay Maini
Samay Mainivor 6 Monaten

let’s archive the world for embodied intelligence!

Profilbild von Kai
Kaivor 6 Monaten

This stack is wild, egocentric video plus tactile plus IMU in one dataset is exactly what makes sim to real less brittle. Fine tuning policies from this should beat pure vision only sets.

Profilbild von Maahir Sachdev
Maahir Sachdevvor 6 Monaten

congrats champ

Profilbild von Kyriakos
Kyriakosvor 6 Monaten

AI needs real world datasets

Profilbild von Manideep
Manideepvor 6 Monaten

this is goated startup in this batch

Profilbild von Jai Bhatia
Jai Bhatiavor 6 Monaten

Congrats boys!

Profilbild von Skyler Chan
Skyler Chanvor 6 Monaten

congrats raj!

Profilbild von BTCaveman.
BTCaveman.vor 6 Monaten

Lfg boys! Kill it ! Proud to see my Indians together like this

Profilbild von The Humanoid Hub
The Humanoid Hubvor 6 Monaten

Models are hungry!!

Profilbild von Mohd Tanveer
Mohd Tanveervor 6 Monaten

Exciting step forward for embodied AI and robotics. Congrats to the team! 🚀

Profilbild von Chain Alpha
Chain Alphavor 6 Monaten

Congrats!

Profilbild von Rushil Agarwal
Rushil Agarwalvor 6 Monaten

Super excited

Profilbild von sreekar nagulapalli
sreekar nagulapallivor 6 Monaten

let’s goooo

Profilbild von Rithvik
Rithvikvor 6 Monaten

Goats

Profilbild von Sanjay Adhikesaven
Sanjay Adhikesavenvor 6 Monaten

congrats!

Profilbild von Abhishek Vadapalli
Abhishek Vadapallivor 6 Monaten

Congrats @babugi28!!

Profilbild von parag parikh
parag parikhvor 6 Monaten

Congratulations and Goodluck ...

Profilbild von Chawit
Chawitvor 6 Monaten

looks amazing!

Profilbild von vishnu manoj
vishnu manojvor 6 Monaten

lfgg🚀

Profilbild von Testimony | Product Designer
Testimony | Product Designervor 6 Monaten

put enough sensors on a human doing manual labor and suddenly you have the most expensive dishwashing video ever recorded. congrats on the launch

Profilbild von justin wang
justin wangvor 6 Monaten

congrats on the launch raj!!

Profilbild von Shamim Hossain
Shamim Hossainvor 6 Monaten

Congratulations on the launch!

Profilbild von interstellar
interstellarvor 6 Monaten

Great !

Profilbild von Mohamed Anis
Mohamed Anisvor 6 Monaten

Excited to see how this enables the next generation of foundation models for physical AI.

Profilbild von The Daily_Ai
The Daily_Aivor 6 Monaten

Exciting stuff! The integration of all those data types will really push robotics forward. Can't wait to see what comes next!

Profilbild von Rayan
Rayanvor 6 Monaten

Congrats !

Profilbild von Ali Taha Brown
Ali Taha Brownvor 6 Monaten

Congrats team

Profilbild von Neel Sharma
Neel Sharmavor 6 Monaten

BANG

Profilbild von Keshav Agarwal
Keshav Agarwalvor 6 Monaten

We have build some apps that we would soon be launching on X and would YC to have a look if they find interesting

Profilbild von Veer Sahasi
Veer Sahasivor 6 Monaten

Awesome!!

Profilbild von Shrinandan Narayanan
Shrinandan Narayananvor 6 Monaten

Congrats guys!

Ähnliche Videos

🔥 JUST IN: Open-source robotics dataset from 100% real-world scenarios! 🤯 Chinese robotics company AGIBOT just released AGIBOT WORLD 2026, an open-source dataset systematically covering key embodied AI research directions. Built entirely from real-world environments: commercial spaces, and homes. Collected using AGIBOT G2 robots in free-form collection mode, providing structured, accurately annotated, high-quality data. Digital twin technology creates 1:1 scale replicas in simulation matching the real environments. Both real-world and simulation data are open-sourced. The AGIBOT G2 platform collects multiple data types simultaneously: RGB(D) cameras, tactile sensors, force sensors, LiDAR, IMU, and full-body joint states. Whole-body control coordinates arms, waist, and hands for complex tasks. First-person teleoperation lets operators control the robot from its perspective. The tasks covered are fine-grained manipulation, ultra-long-horizon tasks, spatial navigation, dual-arm coordination, and multi-agent/human-robot collaboration. The dataset includes error-recovery trajectories with annotations. Most datasets only show successful demonstrations. AGIBOT includes failures and how the robot recovers, teaching models how to handle mistakes. After collection, data is tested through policy training and real-robot deployment to ensure quality. Then processed through industrial quality control with multiple screening and cleaning rounds. Making it open-source accelerates embodied AI research by giving researchers access to high-quality real-world robot data at scale. 🇨🇳 Learn more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

40,583 Aufrufe • vor 5 Monaten

🚨 BREAKING: Microsoft's first robotics foundation model! 🤯 Microsoft just announced Rho-alpha (ρα), their first robotics model derived from the Phi series of vision-language models. Rho-alpha translates natural language commands into control signals for robotic systems performing bimanual manipulation tasks. Commands like "push the green button with the right gripper," "pull out the red wire," "flip the top switch on," or "turn the knob to position 5" get executed directly by dual-arm robots. What makes this different from standard vision-language-action (VLA) models is the additional modalities. Rho-alpha is a VLA+ model that adds tactile sensing to the perceptual mix, with plans to incorporate force feedback. On the learning side, the model is designed to continually improve during deployment by learning from human feedback. The training approach combines trajectories from physical demonstrations and simulated tasks with web-scale visual question answering data. Since teleoperation data is scarce and expensive, Microsoft is using NVIDIA Isaac Sim on Azure to generate physically accurate synthetic datasets via reinforcement learning. These simulated trajectories get combined with commercial and open physical demonstration datasets. The model is currently under evaluation on dual-arm setups and humanoid robots. Microsoft is opening an Early Access Program for organizations interested in evaluating Rho-alpha. Robots that can adapt to dynamic situations and human preferences are more useful in real environments and more trusted by the people operating them. Read more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

60,985 Aufrufe • vor 7 Monaten

JUST IN: Dyna Robotics just published one of the most important research papers in robotics this year. It could fundamentally change how robot foundation models are trained. A scaling law that transfers from human video to robot performance. Dyna-2 is out and it's 🔥 Here's what that means in plain terms. Dyna-2 was pre-trained on ONE MILLION hours of egocentric human video, 170 years of continuous human experience, cooking, folding, assembling, cleaning. And as that human data scaled, robot performance improved. Predictably. Monotonically. Across 39 tasks on two different robot embodiments the model had never seen. → 1,000 hours pre-training → 20% normalised task performance → 10,000 hours → 28% → 100,000 hours → 45% → 1,000,000 hours → 53% Human video exists at effectively unlimited scale. Every cook, every factory worker, every craftsperson wearing a camera is generating training data for future robots. But the finding that stunned even the researchers, world modeling is what makes the transfer work. A model trained to predict future video AND actions massively outperforms one trained on actions alone. Video is the new scaling axis for robotics. One more jaw-dropping data point. 13 minutes of teleoperation data was enough to fine-tune Dyna-2 to open a bottle cap using two five-fingered robot hands. The robots are coming, and they're learning from us directly :D Read more here: Congrats Jason Ma and team! ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

23,576 Aufrufe • vor 1 Monat