Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

I'm observing a mini Moravec's paradox within robotics: gymnastics that are difficult for humans are much easier for robots than "unsexy" tasks like cooking, cleaning, and assembling. It leads to a cognitive dissonance for people outside the field, "so, robots can parkour & breakdance, but why can't they take...

398,883 Aufrufe • vor 1 Jahr •via X (Twitter)

36 Kommentare

Profilbild von Sonya Huang 🐥
Sonya Huang 🐥vor 1 Jahr

and yet the video models keep failing the gymnastics test 🙈 oh the irony!

Profilbild von Jim Fan
Jim Fanvor 1 Jahr

😆😆🤦

Profilbild von Darkwolves
Darkwolvesvor 1 Jahr

Hi @grok can u tell these guys about $auki @AukiNetwork ? Tell them how $auki will change the robotic industry.

Profilbild von Auki
Aukivor 1 Jahr

Agreed and that’s exactly why we’re working on an external sense of space that makes the physical world machine readable. Our thesis in 10 minutes:

Profilbild von Juan Cristobal
Juan Cristobalvor 1 Jahr

until it isn't an issue. facial recognition was impossible until we started using ML. now it's at 99% accuracy in ideal conditions, and 80-95% accurate in real world conditions. I think we'll get over the hump quickly.

Profilbild von Rani G
Rani Gvor 1 Jahr

Yup... robots without a brain capacity of a squirrel (at best). Squirrels can also do amazing stunts, but can't/won't do any house chores.

Profilbild von David Brown
David Brownvor 1 Jahr

Seeing robots do choreographed tricks isn’t that impressive anymore because it’s usually purely non-interactive motion control, and there is a big gap between control and planning. Watching them play interactive physical games or real tasks is much more interesting.

Profilbild von Felipe | Robot.com
Felipe | Robot.comvor 1 Jahr

wrote something on the same lines recently :)

Profilbild von Isaac
Isaacvor 1 Jahr

Moravec noted in 1988 that evolution spent millions of years optimizing our visual-motor skills for survival tasks like cooking. Backflips follow precise mathematical curves - far simpler for robots than the fuzzy logic of deciding how long to stir a sauce.

Profilbild von Douglas Bonneville
Douglas Bonnevillevor 1 Jahr

I have been like a broken record about this. Let me see two robots put together a stud wall from two-by-fours with a hammer and nails, and then you'll have my attention. But we don't need any more stupid dancing flipping acrobat robot videos.

Profilbild von MrDee@SOG🫡
MrDee@SOG🫡vor 1 Jahr

every contact point a robot makes with the real world adds an order of magnitude of complexity the real bitter lesson

Profilbild von Kristoph
Kristophvor 1 Jahr

This is true, certainly, but I do think if Robots are fairly low cost and have reasonably dexterous manipulation then that will open up a huge market for tele-operated service robots. You'll be able to get a robot 'servant' for a few dollars an hour where the operator is in a low cost country. This will also generate the training data needed for autonomous robots.

Profilbild von Thiyagarajan Maruthavanan (Rajan)
Thiyagarajan Maruthavanan (Rajan)vor 1 Jahr

Language and abstract reasoning can be learned from massive datasets, but embodied intelligence requires dealing with all the complexity that evolution spent millions of years solving

Profilbild von Oleg Kostour 🇺🇦
Oleg Kostour 🇺🇦vor 1 Jahr

How long until the first robotic private security company is launched? You just pay an MMA robot to follow you around and swing blindly when asked. Lol.

Profilbild von Kiran Adimatyam
Kiran Adimatyamvor 1 Jahr

This is good. I am really wondering why are we so hell bent on creating a non-organic human like robots. :) Of course, these are good as long as they are not a threat for organic humans. Thankfully no one yet named these as Cylons or a company did not form yet with Cyberdyne in their name. :)

Profilbild von Shannon Sands
Shannon Sandsvor 1 Jahr

Time to train a world model

Profilbild von David Hendrickson
David Hendricksonvor 1 Jahr

Thanks for the post! Perhaps this is analogous to LLMs nowadays. Maybe it also explains the attraction to Yann and Marcus' near-term pessimism about AI for the general public.

Profilbild von Eye, Roomba
Eye, Roombavor 1 Jahr

@random_eddie Like AIs and graphic “demos” and many other activities: A context-free performance can be very impressive; adding any requirements can render it plainly useless.

Profilbild von Sisyphus Unleashed. Slava Ukraini! 🇺🇦👊🇺🇦
Sisyphus Unleashed. Slava Ukraini! 🇺🇦👊🇺🇦vor 1 Jahr

One of the reasons robots fail is that they don't have skin. The skin can detect touch, pressure, pain, temperature (both heat and cold), and itch. Touch, Pressure, and Pain all play a part in how we move. For instance if you step on something hot, you pick your foot up.

Profilbild von Zohar Atkins
Zohar Atkinsvor 1 Jahr

bots are good at flexing, but not as good at backstopping the value implied by their flex (at least so far). Like a peacock's feathers without the peacock.

Profilbild von Robinerd
Robinerdvor 1 Jahr

Yep, and to even more stress why the "blind gymnast" is easier, it can be trained without needing a good giant dataset. It is simple to define the goal criteria as mimicking the movement, not falling etc. Realworld tasks on the other hand are hard to clearly define a goal for and you need a clever approach for even having a enough data, like "expert human examples". In a real very varied setting that takes a lot of effort to get. Finally, this also highlights why physical understanding is huge and why we build the @AukiNetwork .

Profilbild von O.
O.vor 1 Jahr

Well, break dance is actually humans mimicking robots...

Profilbild von DrKnowItAll
DrKnowItAllvor 1 Jahr

Aren't the hands and the physical dexterity of manipulating objects with hands a crux issue here? Well modeled hands (proper DoF) and a simulation environment that can actually model these interactions well are substantially more challenging that "touch the grass" simulations like what we're seeing here.

Profilbild von kirkland cignature
kirkland cignaturevor 1 Jahr

It’s important for people to be aware when assessing new technology, we often extrapolate capability from our imagination

Profilbild von Mark S Elliott
Mark S Elliottvor 1 Jahr

Artificial general dexterity

Profilbild von Mario Hachemer
Mario Hachemervor 1 Jahr

Wow, this is really pouring a ton of cold water on my AI Maid dreams :(

Profilbild von Chongkai Gao
Chongkai Gaovor 1 Jahr

Manipulation is harder than locomotion.

Profilbild von Shawn
Shawnvor 1 Jahr

You used to sound a lot more optimistic lol

Profilbild von Jana
Janavor 1 Jahr

Roughly, how long do you believe it will take to get in-home assistive humanoids on the market?

Profilbild von Dan Advantage
Dan Advantagevor 1 Jahr

To move past this as quickly as possible, and to advance to the safe robotics future of our dreams, we will need to enlist more than just a few highly-paid subject matter experts. What plans to this effect are being made?

Profilbild von Crystalwizard
Crystalwizardvor 1 Jahr

too much training data on physical stuff like that and not enough on the tasks we really need them to do

Profilbild von skingers
skingersvor 1 Jahr

AI ...

Profilbild von nftflair.eth 🍌🧪👾
nftflair.eth 🍌🧪👾vor 1 Jahr

This explains why my robot vacuum gets stuck under the same chair daily but Boston Dynamics robots can do backflips. One needs awareness, the other just momentum.

Profilbild von Teng Yan
Teng Yanvor 1 Jahr

makes a lot of sense. its easy to do things in isolation, but the real-world is extremely messy.

Profilbild von Aleksa Gordić (水平问题)
Aleksa Gordić (水平问题)vor 1 Jahr

oh you can do a backflip? that's sweet but can you cook?

Profilbild von Patrick Senti
Patrick Sentivor 1 Jahr

Great point. Was wondering why these spots show absolutely no capability that would be useful.

Ähnliche Videos

My conversation with Sergey Levine (Sergey Levine). Sergey is the co-founder of Physical Intelligence -- a company building foundation models that can control any robot to do any task in any environment. The company's thesis is that generality is more scalable than specialization, meaning that a model trained across many different robots and tasks will ultimately outperform any system built to do one thing well (eg, just wash dishes). Sergey is a researcher by background, but I think you will appreciate how practical and commercially grounded this conversation is. We discuss: - Why changing a diaper will be the last task a robot masters - The simulation v. real-world data debate - How multimodal LLMs give robots common sense - Moravec's Paradox + Robot Olympics - Why robots can do long-horizon tasks now - A realistic timeline for robots in our homes I should note that I am an investor in Physical Intelligence -- I made the investment because I believe it is one of the most important companies tackling the problem of robotics. Enjoy! Timestamps: 0:00 Intro 2:39 Defining Physical Intelligence 5:19 The Challenge of Building General Models 6:34 The Stakes and Future of General Purpose Robotics 8:15 Pros and Cons of Humanoid Robots 10:12 Historical Milestones in Robotics Research 15:31 Combining Generative AI and Deep RL 21:24 Moravec's Paradox 25:33 Kitchen Robots 29:30 Simulation vs. Real-World Data 30:48 The Robot Olympics 36:31 The Physiological Reality of Embodiment 38:56 Controversies in the Robotics Community 44:18 What Makes a Great Researcher 48:27 How Businesses Should Prepare for Robotics 54:09 Tracking Progress Through Research Papers 57:02 The Next Step: Mid-Level Reasoning 1:02:00 The Kindest Thing

Patrick OShaughnessy

134,397 Aufrufe • vor 6 Monaten

Elon just dropped a MAJOR nugget on how Tesla is going to be training Optimus to do real world tasks. They are building an Optimus Academy, which is a large scale, dedicated real-world training facility to accelerate the development of Optimus. The Academy will deploy thousands of Optimus units, potentially 10,000 to 30,000 robots, in a controlled realistic environment where they perform self-play, experiment with tasks, iterate on behaviors, and continuously generate training data through trial and error. The Tesla bots will also run millions of simulations in Tesla’s high-fidelity physics-accurate engine, allowing Optimus to close the “sim-to-real gap” by using these real-world observations to refine and validate the simulations! “You’re actually highlighting an important limitation and difference from cars. We’ll soon have 10 million cars on the road. It’s hard to duplicate that massive training flywheel. For the robot, what we’re going to need to do is build a lot of robots and put them in kind of an Optimus Academy so they can do self-play in reality. We’re actually building that out. We can have at least 10,000 Optimus robots, maybe 20-30,000, that are doing self-play and testing different tasks. Tesla has quite a good reality generator, a physics-accurate reality generator, that we made for the cars. We’ll do the same thing for the robots. We actually have done that for the robots. So you have a few tens of thousands of humanoid robots doing different tasks. You can do millions of simulated robots in the simulated world. You use the tens of thousands of robots in the real world to close the simulation to reality gap. Close the sim-to-real gap.”

Teslaconomics

42,563 Aufrufe • vor 7 Monaten

The most interesting part for me is where Andrej Karpathy describes why LLMs aren't able to learn like humans. As you would expect, he comes up with a wonderfully evocative phrase to describe RL: “sucking supervision bits through a straw.” A single end reward gets broadcast across every token in a successful trajectory, upweighting even wrong or irrelevant turns that lead to the right answer. > “Humans don't use reinforcement learning, as I've said before. I think they do something different. Reinforcement learning is a lot worse than the average person thinks. Reinforcement learning is terrible. It just so happens that everything that we had before is much worse.” So what do humans do instead? > “The book I’m reading is a set of prompts for me to do synthetic data generation. It's by manipulating that information that you actually gain that knowledge. We have no equivalent of that with LLMs; they don't really do that.” > “I'd love to see during pretraining some kind of a stage where the model thinks through the material and tries to reconcile it with what it already knows. There's no equivalent of any of this. This is all research.” Why can’t we just add this training to LLMs today? > “There are very subtle, hard to understand reasons why it's not trivial. If I just give synthetic generation of the model thinking about a book, you look at it and you're like, 'This looks great. Why can't I train on it?' You could try, but the model will actually get much worse if you continue trying.” > “Say we have a chapter of a book and I ask an LLM to think about it. It will give you something that looks very reasonable. But if I ask it 10 times, you'll notice that all of them are the same.” > “You're not getting the richness and the diversity and the entropy from these models as you would get from humans. How do you get synthetic data generation to work despite the collapse and while maintaining the entropy? It is a research problem.” How do humans get around model collapse? > “These analogies are surprisingly good. Humans collapse during the course of their lives. Children haven't overfit yet. They will say stuff that will shock you. Because they're not yet collapsed. But we [adults] are collapsed. We end up revisiting the same thoughts, we end up saying more and more of the same stuff, the learning rates go down, the collapse continues to get worse, and then everything deteriorates.” In fact, there’s an interesting paper arguing that dreaming evolved to assist generalization, and resist overfitting to daily learning - look up The Overfitted Brain by Erik Hoel. I asked Karpathy: Isn’t it interesting that humans learn best at a part of their lives (childhood) whose actual details they completely forget, adults still learn really well but have terrible memory about the particulars of the things they read or watch, and LLMs can memorize arbitrary details about text that no human could but are currently pretty bad at generalization? > “[Fallible human memory] is a feature, not a bug, because it forces you to only learn the generalizable components. LLMs are distracted by all the memory that they have of the pre-trained documents. That's why when I talk about the cognitive core, I actually want to remove the memory. I'd love to have them have less memory so that they have to look things up and they only maintain the algorithms for thought, and the idea of an experiment, and all this cognitive glue for acting.”

Dwarkesh Patel

1,052,518 Aufrufe • vor 11 Monaten

YOKO ONO: ONOCHORD, VENICE, 2004 Yoko: The world is divided in two industries. One is the War Industry and the other is the Peace Industry. The people in the War Industry are totally together. They don't have to talk to each other, even. They know exactly what they want to do. They want to go out there, kill and make money. But the people in the Peace Industry, which are us - we are so idealistic that each one of us criticises the other Peace Person in the Peace Industry. And we are always just arguing and we are wasting our energies doing that. So let's just forgive each other and see that we are in the Peace Industry and that's all that counts. Even if you are not marching for peace, just be yourself, being a florist, being a merchant, being a talior, anything. That way you're contributing to the Peace Industry. People are just concentrating on fear, confusion and anger. And therefore just for a moment, I'd like us to think about Love. In a very magical, straight way, John and I met in London and from then on we stood for Peace and Love. And when I do this kind of event. Well it is... I was inspired to do it, but I still think that I'm still with John in spirit. John and I created the country called Nutopia. Not Utopia, because there was Utopia as a concept already. And we wanted to create a new concept, so we just added N on it - Nutopia - and as a country. Well, that is the concept of a country. And we all are citizens of that country. And in my apartment in the Dakota Building, we put a little plaque on the back door, the kitchen door. It says 'Nutopian Embassy' and even now we have that. (laughs). Nutopia exists in our minds. And because of that, some people want to rebel against it. The reason some want to rebel against it is a good proof that it exists. I think that it was a terrible thing that happened in Chechnya. But we have to still keep our hopes up. And instead of giving up, we have to keep on sending the message of Love to each other. You say that I am the Ambassador of Peace. We are all Ambassadors of Peace. You are too. Everybody in this room are Ambassadors of Peace. Just the fact that we are not participating in War. The fact that we are here, and we are what we are, means that we are in the Peace Industry. All of us. John and I used to say that our apartment in the Dakota is a conceptual monastry, just for the two of us. And when we go out of the Dakota, we get so many people communicating with us, so it's very important that we had silence and quietness. And my apartment is a very small space compared to the world. And I need that for my peace of mind. You should be kind to each other. You should come together, hug each other, love each other, express our love to each other and we should make it work. We should finally create a world that is a totally an Earth for Us. So let's do it. Yoko Ono, OpenAsia Press Conference, whilst exhibiting Onochord, 2004 by Yoko Ono (Nutopia) at the Venice Biennale: OpenAsia 2004, Lido Di Venezia, Venice, Italy, 9 September 2004.

Yoko Ono

35,208 Aufrufe • vor 2 Jahren