Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Human Archive builds datasets to model sensorimotor intelligence for robotics and world models. HA-Multi is the largest dataset of manual labor tasks aligning egocentric RGB, stereo depth (active IR), tactile gloves, body IMUs, and wrist cameras. Congrats on the launch, Raj Patel, Shloke Patel, Samay Maini, and Rushil Agarwal!

110,791 görüntüleme • 6 ay önce •via X (Twitter)

34 Yorum

Runtime profil fotoğrafı
Runtime6 ay önce

Onwards and upwards!

Panda Hodler (YC Intern) profil fotoğrafı
Panda Hodler (YC Intern)6 ay önce

This seems to equip robots with "human-level tactile and visual nerves." The team's approach is very interesting i think😆

WildPinesAI profil fotoğrafı
WildPinesAI6 ay önce

tactile force maps aligned with vision is the missing dataset every robotics foundation model is starving for

Samay Maini profil fotoğrafı
Samay Maini6 ay önce

let’s archive the world for embodied intelligence!

Kai profil fotoğrafı
Kai6 ay önce

This stack is wild, egocentric video plus tactile plus IMU in one dataset is exactly what makes sim to real less brittle. Fine tuning policies from this should beat pure vision only sets.

Maahir Sachdev profil fotoğrafı
Maahir Sachdev6 ay önce

congrats champ

Kyriakos profil fotoğrafı
Kyriakos6 ay önce

AI needs real world datasets

Manideep profil fotoğrafı
Manideep6 ay önce

this is goated startup in this batch

Jai Bhatia profil fotoğrafı
Jai Bhatia6 ay önce

Congrats boys!

Skyler Chan profil fotoğrafı
Skyler Chan6 ay önce

congrats raj!

BTCaveman. profil fotoğrafı
BTCaveman.6 ay önce

Lfg boys! Kill it ! Proud to see my Indians together like this

The Humanoid Hub profil fotoğrafı
The Humanoid Hub6 ay önce

Models are hungry!!

Mohd Tanveer profil fotoğrafı
Mohd Tanveer6 ay önce

Exciting step forward for embodied AI and robotics. Congrats to the team! 🚀

Chain Alpha profil fotoğrafı
Chain Alpha6 ay önce

Congrats!

Rushil Agarwal profil fotoğrafı
Rushil Agarwal6 ay önce

Super excited

sreekar nagulapalli profil fotoğrafı
sreekar nagulapalli6 ay önce

let’s goooo

Rithvik profil fotoğrafı
Rithvik6 ay önce

Goats

Sanjay Adhikesaven profil fotoğrafı
Sanjay Adhikesaven6 ay önce

congrats!

Abhishek Vadapalli profil fotoğrafı
Abhishek Vadapalli6 ay önce

Congrats @babugi28!!

parag parikh profil fotoğrafı
parag parikh6 ay önce

Congratulations and Goodluck ...

Chawit profil fotoğrafı
Chawit6 ay önce

looks amazing!

vishnu manoj profil fotoğrafı
vishnu manoj6 ay önce

lfgg🚀

Testimony | Product Designer profil fotoğrafı
Testimony | Product Designer6 ay önce

put enough sensors on a human doing manual labor and suddenly you have the most expensive dishwashing video ever recorded. congrats on the launch

justin wang profil fotoğrafı
justin wang6 ay önce

congrats on the launch raj!!

Shamim Hossain profil fotoğrafı
Shamim Hossain6 ay önce

Congratulations on the launch!

interstellar profil fotoğrafı
interstellar6 ay önce

Great !

Mohamed Anis profil fotoğrafı
Mohamed Anis6 ay önce

Excited to see how this enables the next generation of foundation models for physical AI.

The Daily_Ai profil fotoğrafı
The Daily_Ai6 ay önce

Exciting stuff! The integration of all those data types will really push robotics forward. Can't wait to see what comes next!

Rayan profil fotoğrafı
Rayan6 ay önce

Congrats !

Ali Taha Brown profil fotoğrafı
Ali Taha Brown6 ay önce

Congrats team

Neel Sharma profil fotoğrafı
Neel Sharma6 ay önce

BANG

Keshav Agarwal profil fotoğrafı
Keshav Agarwal6 ay önce

We have build some apps that we would soon be launching on X and would YC to have a look if they find interesting

Veer Sahasi profil fotoğrafı
Veer Sahasi6 ay önce

Awesome!!

Shrinandan Narayanan profil fotoğrafı
Shrinandan Narayanan6 ay önce

Congrats guys!

Benzer Videolar

🔥 JUST IN: Open-source robotics dataset from 100% real-world scenarios! 🤯 Chinese robotics company AGIBOT just released AGIBOT WORLD 2026, an open-source dataset systematically covering key embodied AI research directions. Built entirely from real-world environments: commercial spaces, and homes. Collected using AGIBOT G2 robots in free-form collection mode, providing structured, accurately annotated, high-quality data. Digital twin technology creates 1:1 scale replicas in simulation matching the real environments. Both real-world and simulation data are open-sourced. The AGIBOT G2 platform collects multiple data types simultaneously: RGB(D) cameras, tactile sensors, force sensors, LiDAR, IMU, and full-body joint states. Whole-body control coordinates arms, waist, and hands for complex tasks. First-person teleoperation lets operators control the robot from its perspective. The tasks covered are fine-grained manipulation, ultra-long-horizon tasks, spatial navigation, dual-arm coordination, and multi-agent/human-robot collaboration. The dataset includes error-recovery trajectories with annotations. Most datasets only show successful demonstrations. AGIBOT includes failures and how the robot recovers, teaching models how to handle mistakes. After collection, data is tested through policy training and real-robot deployment to ensure quality. Then processed through industrial quality control with multiple screening and cleaning rounds. Making it open-source accelerates embodied AI research by giving researchers access to high-quality real-world robot data at scale. 🇨🇳 Learn more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

40,583 görüntüleme • 5 ay önce

🚨 BREAKING: Microsoft's first robotics foundation model! 🤯 Microsoft just announced Rho-alpha (ρα), their first robotics model derived from the Phi series of vision-language models. Rho-alpha translates natural language commands into control signals for robotic systems performing bimanual manipulation tasks. Commands like "push the green button with the right gripper," "pull out the red wire," "flip the top switch on," or "turn the knob to position 5" get executed directly by dual-arm robots. What makes this different from standard vision-language-action (VLA) models is the additional modalities. Rho-alpha is a VLA+ model that adds tactile sensing to the perceptual mix, with plans to incorporate force feedback. On the learning side, the model is designed to continually improve during deployment by learning from human feedback. The training approach combines trajectories from physical demonstrations and simulated tasks with web-scale visual question answering data. Since teleoperation data is scarce and expensive, Microsoft is using NVIDIA Isaac Sim on Azure to generate physically accurate synthetic datasets via reinforcement learning. These simulated trajectories get combined with commercial and open physical demonstration datasets. The model is currently under evaluation on dual-arm setups and humanoid robots. Microsoft is opening an Early Access Program for organizations interested in evaluating Rho-alpha. Robots that can adapt to dynamic situations and human preferences are more useful in real environments and more trusted by the people operating them. Read more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

60,985 görüntüleme • 7 ay önce

JUST IN: Dyna Robotics just published one of the most important research papers in robotics this year. It could fundamentally change how robot foundation models are trained. A scaling law that transfers from human video to robot performance. Dyna-2 is out and it's 🔥 Here's what that means in plain terms. Dyna-2 was pre-trained on ONE MILLION hours of egocentric human video, 170 years of continuous human experience, cooking, folding, assembling, cleaning. And as that human data scaled, robot performance improved. Predictably. Monotonically. Across 39 tasks on two different robot embodiments the model had never seen. → 1,000 hours pre-training → 20% normalised task performance → 10,000 hours → 28% → 100,000 hours → 45% → 1,000,000 hours → 53% Human video exists at effectively unlimited scale. Every cook, every factory worker, every craftsperson wearing a camera is generating training data for future robots. But the finding that stunned even the researchers, world modeling is what makes the transfer work. A model trained to predict future video AND actions massively outperforms one trained on actions alone. Video is the new scaling axis for robotics. One more jaw-dropping data point. 13 minutes of teleoperation data was enough to fine-tune Dyna-2 to open a bottle cap using two five-fingered robot hands. The robots are coming, and they're learning from us directly :D Read more here: Congrats Jason Ma and team! ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

23,576 görüntüleme • 1 ay önce