Загрузка видео...

Не удалось загрузить видео

На главную

Human Archive builds datasets to model sensorimotor intelligence for robotics and world models. HA-Multi is the largest dataset of manual labor tasks aligning egocentric RGB, stereo depth (active IR), tactile gloves, body IMUs, and wrist cameras. Congrats on the launch, Raj Patel, Shloke Patel, Samay Maini, and Rushil Agarwal!

110,791 просмотров • 6 месяцев назад •via X (Twitter)

Комментарии: 34

Фото профиля Runtime
Runtime6 месяцев назад

Onwards and upwards!

Фото профиля Panda Hodler (YC Intern)
Panda Hodler (YC Intern)6 месяцев назад

This seems to equip robots with "human-level tactile and visual nerves." The team's approach is very interesting i think😆

Фото профиля WildPinesAI
WildPinesAI6 месяцев назад

tactile force maps aligned with vision is the missing dataset every robotics foundation model is starving for

Фото профиля Samay Maini
Samay Maini6 месяцев назад

let’s archive the world for embodied intelligence!

Фото профиля Kai
Kai6 месяцев назад

This stack is wild, egocentric video plus tactile plus IMU in one dataset is exactly what makes sim to real less brittle. Fine tuning policies from this should beat pure vision only sets.

Фото профиля Maahir Sachdev
Maahir Sachdev6 месяцев назад

congrats champ

Фото профиля Kyriakos
Kyriakos6 месяцев назад

AI needs real world datasets

Фото профиля Manideep
Manideep6 месяцев назад

this is goated startup in this batch

Фото профиля Jai Bhatia
Jai Bhatia6 месяцев назад

Congrats boys!

Фото профиля Skyler Chan
Skyler Chan6 месяцев назад

congrats raj!

Фото профиля BTCaveman.
BTCaveman.6 месяцев назад

Lfg boys! Kill it ! Proud to see my Indians together like this

Фото профиля The Humanoid Hub
The Humanoid Hub6 месяцев назад

Models are hungry!!

Фото профиля Mohd Tanveer
Mohd Tanveer6 месяцев назад

Exciting step forward for embodied AI and robotics. Congrats to the team! 🚀

Фото профиля Chain Alpha
Chain Alpha6 месяцев назад

Congrats!

Фото профиля Rushil Agarwal
Rushil Agarwal6 месяцев назад

Super excited

Фото профиля sreekar nagulapalli
sreekar nagulapalli6 месяцев назад

let’s goooo

Фото профиля Rithvik
Rithvik6 месяцев назад

Goats

Фото профиля Sanjay Adhikesaven
Sanjay Adhikesaven6 месяцев назад

congrats!

Фото профиля Abhishek Vadapalli
Abhishek Vadapalli6 месяцев назад

Congrats @babugi28!!

Фото профиля parag parikh
parag parikh6 месяцев назад

Congratulations and Goodluck ...

Фото профиля Chawit
Chawit6 месяцев назад

looks amazing!

Фото профиля vishnu manoj
vishnu manoj6 месяцев назад

lfgg🚀

Фото профиля Testimony | Product Designer
Testimony | Product Designer6 месяцев назад

put enough sensors on a human doing manual labor and suddenly you have the most expensive dishwashing video ever recorded. congrats on the launch

Фото профиля justin wang
justin wang6 месяцев назад

congrats on the launch raj!!

Фото профиля Shamim Hossain
Shamim Hossain6 месяцев назад

Congratulations on the launch!

Фото профиля interstellar
interstellar6 месяцев назад

Great !

Фото профиля Mohamed Anis
Mohamed Anis6 месяцев назад

Excited to see how this enables the next generation of foundation models for physical AI.

Фото профиля The Daily_Ai
The Daily_Ai6 месяцев назад

Exciting stuff! The integration of all those data types will really push robotics forward. Can't wait to see what comes next!

Фото профиля Rayan
Rayan6 месяцев назад

Congrats !

Фото профиля Ali Taha Brown
Ali Taha Brown6 месяцев назад

Congrats team

Фото профиля Neel Sharma
Neel Sharma6 месяцев назад

BANG

Фото профиля Keshav Agarwal
Keshav Agarwal6 месяцев назад

We have build some apps that we would soon be launching on X and would YC to have a look if they find interesting

Фото профиля Veer Sahasi
Veer Sahasi6 месяцев назад

Awesome!!

Фото профиля Shrinandan Narayanan
Shrinandan Narayanan6 месяцев назад

Congrats guys!

Похожие видео

🔥 JUST IN: Open-source robotics dataset from 100% real-world scenarios! 🤯 Chinese robotics company AGIBOT just released AGIBOT WORLD 2026, an open-source dataset systematically covering key embodied AI research directions. Built entirely from real-world environments: commercial spaces, and homes. Collected using AGIBOT G2 robots in free-form collection mode, providing structured, accurately annotated, high-quality data. Digital twin technology creates 1:1 scale replicas in simulation matching the real environments. Both real-world and simulation data are open-sourced. The AGIBOT G2 platform collects multiple data types simultaneously: RGB(D) cameras, tactile sensors, force sensors, LiDAR, IMU, and full-body joint states. Whole-body control coordinates arms, waist, and hands for complex tasks. First-person teleoperation lets operators control the robot from its perspective. The tasks covered are fine-grained manipulation, ultra-long-horizon tasks, spatial navigation, dual-arm coordination, and multi-agent/human-robot collaboration. The dataset includes error-recovery trajectories with annotations. Most datasets only show successful demonstrations. AGIBOT includes failures and how the robot recovers, teaching models how to handle mistakes. After collection, data is tested through policy training and real-robot deployment to ensure quality. Then processed through industrial quality control with multiple screening and cleaning rounds. Making it open-source accelerates embodied AI research by giving researchers access to high-quality real-world robot data at scale. 🇨🇳 Learn more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

40,583 просмотров • 5 месяцев назад

🚨 BREAKING: Microsoft's first robotics foundation model! 🤯 Microsoft just announced Rho-alpha (ρα), their first robotics model derived from the Phi series of vision-language models. Rho-alpha translates natural language commands into control signals for robotic systems performing bimanual manipulation tasks. Commands like "push the green button with the right gripper," "pull out the red wire," "flip the top switch on," or "turn the knob to position 5" get executed directly by dual-arm robots. What makes this different from standard vision-language-action (VLA) models is the additional modalities. Rho-alpha is a VLA+ model that adds tactile sensing to the perceptual mix, with plans to incorporate force feedback. On the learning side, the model is designed to continually improve during deployment by learning from human feedback. The training approach combines trajectories from physical demonstrations and simulated tasks with web-scale visual question answering data. Since teleoperation data is scarce and expensive, Microsoft is using NVIDIA Isaac Sim on Azure to generate physically accurate synthetic datasets via reinforcement learning. These simulated trajectories get combined with commercial and open physical demonstration datasets. The model is currently under evaluation on dual-arm setups and humanoid robots. Microsoft is opening an Early Access Program for organizations interested in evaluating Rho-alpha. Robots that can adapt to dynamic situations and human preferences are more useful in real environments and more trusted by the people operating them. Read more here: ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

60,985 просмотров • 7 месяцев назад

JUST IN: Dyna Robotics just published one of the most important research papers in robotics this year. It could fundamentally change how robot foundation models are trained. A scaling law that transfers from human video to robot performance. Dyna-2 is out and it's 🔥 Here's what that means in plain terms. Dyna-2 was pre-trained on ONE MILLION hours of egocentric human video, 170 years of continuous human experience, cooking, folding, assembling, cleaning. And as that human data scaled, robot performance improved. Predictably. Monotonically. Across 39 tasks on two different robot embodiments the model had never seen. → 1,000 hours pre-training → 20% normalised task performance → 10,000 hours → 28% → 100,000 hours → 45% → 1,000,000 hours → 53% Human video exists at effectively unlimited scale. Every cook, every factory worker, every craftsperson wearing a camera is generating training data for future robots. But the finding that stunned even the researchers, world modeling is what makes the transfer work. A model trained to predict future video AND actions massively outperforms one trained on actions alone. Video is the new scaling axis for robotics. One more jaw-dropping data point. 13 minutes of teleoperation data was enough to fine-tune Dyna-2 to open a bottle cap using two five-fingered robot hands. The robots are coming, and they're learning from us directly :D Read more here: Congrats Jason Ma and team! ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

23,576 просмотров • 1 месяц назад