Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

A video from a Pi robot deployed at Dandelion Chocolate, fully autonomous w/ no interventions. 🤖 Deploying robots has taught us surprising lessons about the gap between proof-of-concept (i.e. building one box) and real-world utility (productively building boxes for hours). I expected that the hard part is building the...

104,016 görüntüleme • 18 gün önce •via X (Twitter)

12 Yorum

ayush profil fotoğrafı
ayush18 gün önce

@physical_int oh so thats what this guy was up to

Will Bright profil fotoğrafı
Will Bright18 gün önce

@physical_int @ASong408 What'd I tell you about CPG?

ethereagle · building profil fotoğrafı
ethereagle · building18 gün önce

@physical_int the clip is zero-intervention, which is the easy number. the one that decides is what failed at minute 40 that minute 2 never saw. was it grasp retry pileup, or perception drift on the chocolate?

tab profil fotoğrafı
tab18 gün önce

@physical_int Seeing it in a real setting like this, applications keep popping into my head.

Jakie PLA profil fotoğrafı
Jakie PLA18 gün önce

@physical_int I could not help watching this on repeat. THE STACK is the hard part. Hours with zero babysitting beats one perfect pick :)

Senthilnathan K profil fotoğrafı
Senthilnathan K18 gün önce

@physical_int The delayed collapse is the part demos hide, a placement can look fine and still be a bad policy because the cost only shows up ten boxes later. It’s more of a leftover-state problem than a dexterity gap.

Arthur Petron profil fotoğrafı
Arthur Petron18 gün önce

@physical_int Awesome.

Porvesh Balasubramanian profil fotoğrafı
Porvesh Balasubramanian18 gün önce

@physical_int Awesome!

Alex Kennberg profil fotoğrafı
Alex Kennberg18 gün önce

@ZeYanjie @physical_int Dandelion Chocolate boxes are nice.

Paige profil fotoğrafı
Paige18 gün önce

@physical_int Would be cool to see fine tuned Pi models in the wild. What other sites are you interested in?

RealMan Robotics profil fotoğrafı
RealMan Robotics18 gün önce

@physical_int This is the real benchmark for deployment: not whether a robot can complete the task once, but whether it can keep doing it reliably for hours with no intervention. That gap between demo and utility is where most of the hard work lives.

Leo Lu profil fotoğrafı
Leo Lu18 gün önce

@physical_int Robotics stops being a demo and starts becoming a product.

Benzer Videolar

We trained a humanoid with 22-DoF dexterous hands to assemble model cars, operate syringes, sort poker cards, fold/roll shirts, all learned primarily from 20,000+ hours of egocentric human video with no robot in the loop. Humans are the most scalable embodiment on the planet. We discovered a near-perfect log-linear scaling law (R² = 0.998) between human video volume and action prediction loss, and this loss directly predicts real-robot success rate. Humanoid robots will be the end game, because they are the practical form factor with minimal embodiment gap from humans. Call it the Bitter Lesson of robot hardware: the kinematic similarity lets us simply retarget human finger motion onto dexterous robot hand joints. No learned embeddings, no fancy transfer algorithms needed. Relative wrist motion + retargeted 22-DoF finger actions serve as a unified action space that carries through from pre-training to robot execution. Our recipe is called "EgoScale": - Pre-train GR00T N1.5 on 20K hours of human video, mid-train with only 4 hours (!) of robot play data with Sharpa hands. 54% gains over training from scratch across 5 highly dexterous tasks. - Most surprising result: a *single* teleop demo is sufficient to learn a never-before-seen task. Our recipe enables extreme data efficiency. - Although we pre-train in 22-DoF hand joint space, the policy transfers to a Unitree G1 with 7-DoF tri-finger hands. 30%+ gains over training on G1 data alone. The scalable path to robot dexterity was never more robots. It was always us. Deep dives in thread:

Jim Fan

301,632 görüntüleme • 7 ay önce

JUST IN: Dyna Robotics just published one of the most important research papers in robotics this year. It could fundamentally change how robot foundation models are trained. A scaling law that transfers from human video to robot performance. Dyna-2 is out and it's 🔥 Here's what that means in plain terms. Dyna-2 was pre-trained on ONE MILLION hours of egocentric human video, 170 years of continuous human experience, cooking, folding, assembling, cleaning. And as that human data scaled, robot performance improved. Predictably. Monotonically. Across 39 tasks on two different robot embodiments the model had never seen. → 1,000 hours pre-training → 20% normalised task performance → 10,000 hours → 28% → 100,000 hours → 45% → 1,000,000 hours → 53% Human video exists at effectively unlimited scale. Every cook, every factory worker, every craftsperson wearing a camera is generating training data for future robots. But the finding that stunned even the researchers, world modeling is what makes the transfer work. A model trained to predict future video AND actions massively outperforms one trained on actions alone. Video is the new scaling axis for robotics. One more jaw-dropping data point. 13 minutes of teleoperation data was enough to fine-tune Dyna-2 to open a bottle cap using two five-fingered robot hands. The robots are coming, and they're learning from us directly :D Read more here: Congrats Jason Ma and team! ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

23,681 görüntüleme • 1 ay önce