Загрузка видео...

Не удалось загрузить видео

На главную

Introduce HumanPlus - Autonomous Skills part Humanoids are born for using human data. Imitating humans, our humanoid learns: - fold sweatshirts - unload objects from warehouse racks - diverse locomotion skills (squatting, jumping, standing) - greet another robot Open-sourced!

158,237 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 10

Фото профиля Zipeng Fu
Zipeng Fu2 лет назад

We build our customized 33-DoF humanoid, and a data collection pipeline through real-time shadowing in the real world.

Фото профиля Zipeng Fu
Zipeng Fu2 лет назад

Using the data collected through shadowing, we then perform supervised behavior cloning to train skill policies using egocentric vision. We introduce Humanoid Imitation Transformer. Based on ACT, HIT adds forward dynamics prediction on image feature space as a regularization.

Фото профиля Zipeng Fu
Zipeng Fu2 лет назад

Compared to baselines, HIT uses - binocular vision, thus having implicit stereos for depth information - visual feedback better, avoiding overfitting to proprioception given small-sized demos

Фото профиля Zipeng Fu
Zipeng Fu2 лет назад

Besides vision-based whole-body manipulation skills, our humanoid has strong locomotion skills: - outperforming H1 default standing controller under strong perturbation forces - enabling more whole-body skills like squatting and jumping

Фото профиля Zipeng Fu
Zipeng Fu2 лет назад

This project is not possible without our team of experts, covering from computer graphics to robot learning to robot hardware: - co-leads: @qingqing_zhao_ @Qi_Wu577 - advisors: @chelseabfinn @GordonWetzstein project website: hardware: code:

Фото профиля Karol Hausman
Karol Hausman2 лет назад

Congrats @zipengfu !! 🎉

Фото профиля Yu Yang
Yu Yang2 лет назад

Pretty cool!

Фото профиля Joe Valentine
Joe Valentine2 лет назад

@DrDisrespect robodoc?

Фото профиля BotBot
BotBot2 лет назад

Wow! Incredible, is the demo unit the Unitree H1?

Фото профиля Ak
Ak2 лет назад

@heyjchu

Похожие видео

Excited to announce GR00T N1, the world’s first open foundation model for humanoid robots! We are on a mission to democratize Physical AI. The power of general robot brain, in the palm of your hand - with only 2B parameters, N1 learns from the most diverse physical action dataset ever compiled and punches above its weight: - Real humanoid teleoperation data. - Large-scale simulation data: we are open-sourcing 300K+ trajectories! - Neural trajectories: we apply SOTA video generation models to “hallucinate” new synthetic data that features accurate physics in pixels. Using Jensen’s words, “systematically infinite data”! - Latent actions: we develop novel algorithms to extract action tokens from in-the-wild human videos and neural generated videos. GR00T N1 is a single end-to-end neural net, from photons to actions: - Vision-Language Model (System 2) that interprets the physical world through vision and language instructions, enabling robots to reason about their environment and instructions, and plan the right actions. - Diffusion Transformer (System 1) that “renders” smooth and precise motor actions at 120 Hz, executing the latent plan made by System 2. We deploy N1 on GR1 robot, 1X Neo robot, and a large collection of simulation benchmarks. N1 achieves up to +30% boost in diverse manipulation tasks for household and industrial settings. While humanoid robots are the main focus of N1, our model also supports cross-embodiment. We finetune it to work on the $110 HuggingFace LeRobot SO100 robot arm! Open robot brain runs on open hardware. Sounds just right. Let’s solve robotics, together, one token at a time. Links to our Whitepaper, Github repo, HuggingFace model, and open dataset page in the thread: 🧵

Jim Fan

466,333 просмотров • 1 год назад

I don’t know if we live in a Matrix, but I know for sure that robots will spend most of their lives in simulation. Let machines train machines. I’m excited to introduce DexMimicGen, a massive-scale synthetic data generator that enables a humanoid robot to learn complex skills from only a handful of human demonstrations. Yes, as few as 5! DexMimicGen addresses the biggest pain point in robotics: where do we get data? Unlike with LLMs, where vast amounts of texts are readily available, you cannot simply download motor control signals from the internet. So researchers teleoperate the robots to collect motion data via XR headsets. They have to repeat the same skill over and over and over again, because neural nets are data hungry. This is a very slow and uncomfortable process. At NVIDIA, we believe the majority of high-quality tokens for robot foundation models will come from simulation. What DexMimicGen does is to trade GPU compute time for human time. It takes one motion trajectory from human, and multiplies into 1000s of new trajectories. A robot brain trained on this augmented dataset will generalize far better in the real world. Think of DexMimicGen as a learning signal amplifier. It maps a small dataset to a large (de facto infinite) dataset, using physics simulation in the loop. In this way, we free humans from babysitting the bots all day. The future of robot data is generative. The future of the entire robot learning pipeline will also be generative. 🧵

Jim Fan

165,246 просмотров • 1 год назад

China unveils humanoid robot with lifelike skin and blinking eyes built for daily life | Prabhat Ranjan Mishra, Interesting Engineering Large Language Models (LLMs) and Vision-Language Models (VLMs) help process and interpret complex data from human interactions. A Shanghai-based company has developed humanoid robots that appear as real as humans. The advanced bionic humanoid robot is integrated with self-supervised AI algorithms. Named Elf V1, the robot can perceive the world, communicate, learn, and interact intelligently with its surroundings. Developed by AheadForm Technology, the robot offers up to 30 degrees of freedom, powered by a precise control system and an advanced AI learning algorithm. Robot offers expressive facial features The robot offers expressive facial features, moving eyes, and synchronized speech. It can also convey emotions and understand human non-verbal cues, making interactions more natural and engaging. The robot has highly interactive capabilities and lifelike appearances. AheadForm expects that its robots could soon seamlessly integrate into daily life, providing assistance, companionship, and support across various industries. “We believe that by developing realistic and expressive robot heads, we can bridge the gap between humans and machines, fostering a new era of interactive and intelligent robotics,” said the company in a statement. Reports revealed that to avoid the “uncanny valley” effect and be able to interact with us, they are given lifelike skin and capabilities to read our emotions and respond appropriately using dynamic expression simulation and emotion generation tech. Bionic skin and high-precision control system The Elf V1 series of humanoids features 30 facial muscles animated by brushless micro-motors and managed by a high-precision control system. Paired with an ability to detect their users’ emotions with low latency and bionic skin, their facial expressions are nearly identical to those of humans, reported CGTN. The company claims it’s pioneering the development of realistic humanoid robots designed to revolutionize human-robot interaction. It’s enhancing sophisticated humanoid robot heads that can express emotions, perceive their environment, and interact seamlessly with humans. By combining cutting-edge AI and advanced robotics, AheadForm aims to bring life to machines and transform how humans engage with technology. AI models boost robots’ responsiveness Seamless integration of Large Language Models (LLMs) and Vision-Language Models (VLMs) into the humanoid robots can help them process and interpret complex data from human interactions, enabling the robot to learn and adapt in real-time, achieving human-level understanding and responsiveness. AheadForm uses Brushless Motors that deliver ultra-quiet operation and high responsiveness, specifically designed for precision facial movements in humanoid robots. With its compact size, lightweight design, and energy efficiency, this motor is the ideal choice for next-generation robots that require precise, subtle facial control to create a truly human-like experience. Previously, the company unveiled the Lan Series that features realistic humanoid robots with soft skin and 10 degrees of freedom, offering a lifelike appearance and intuitive movements. This series is designed for cost-efficiency, for applications prioritizing mobility and manipulation.

Owen Gregorian

179,005 просмотров • 9 месяцев назад

Today, we're joined by Nikita Rudin, co-founder and CEO of Flexion to discuss the gap between current robotic capabilities and what’s required to deploy fully autonomous robots in the real world. Nikita explains how reinforcement learning and simulation have driven rapid progress in robot locomotion—and why locomotion is still far from “solved.” We dig into the sim2real gap, and how adding visual inputs introduces noise and significantly complicates sim-to-real transfer. We also explore the debate between end-to-end models and modular approaches, and why separating locomotion, planning, and semantics remains a pragmatic approach today. Nikita also introduces the concept of "real-to-sim", which uses real-world data to refine simulation parameters for higher fidelity training, discusses how reinforcement learning, imitation learning, and teleoperation data are combined to train robust policies for both quadruped and humanoid robots, and introduces Flexion's hierarchical approach that utilizes pre-trained Vision-Language Models (VLMs) for high-level task orchestration with Vision-Language-Action (VLA) models and low-level whole-body trackers. Finally, Nikita shares the behind-the-scenes in humanoid robot demos, his take on reinforcement learning in simulation versus the real world, the nuances of reward tuning, and offers practical advice for researchers and practitioners looking to get started in robotics today. 🗒️ For the full list of resources for this episode, visit the show notes page: 📖 CHAPTERS =============================== 00:00 - Introduction 04:07 - Is robot locomotion solved? 06:04 - Sim-to-real gap 08:58 - Adding semantics to policies 09:42 - Modular vs end-to-end architectures 10:29 - Planner model 12:21 - Adapting RL techniques from quadrupeds to humanoids 15:39 - Behind robot demos 18:09 - Humanoid robots in home environments 22:03 - Training approach 23:56 - VLA models 27:59 - Closing the sim-to-real gap 32:55 - Task orchestration using VLMs 36:38 - Tool use 38:10 - Model hierarchy 43:37 - Simulator versus simulation environment 44:57 - Combining imitation learning and reinforcement learning 46:42 - RL in real world versus RL in simulation 52:58 - Reward tuning and value functions in robotics 56:38 - Predictions 1:00:10 - Humanoids, quadropeds, and wheeled platforms 1:02:45 - Advice, recommended robot kits, and community pla

The TWIML AI Podcast

22,592 просмотров • 6 месяцев назад

A truly staggering number of AI-powered humanoid robots are emerging from Shenzhen, the booming Chinese tech hub. Shenzhen Dobot, which rose from a Kickstarter campaign to become an industrial automation leader, says its first humanoid robot, Atom, has entered mass production. Atom is positioned as a general-purpose cross-industry worker. Priced around $27,000, Atom boasts fine motor control that’s precise enough to pick up a cherry by the stem. Another major player is Pudu Robotics, which recently shipped its 100,000th service robot. Pudu is commercializing its wheeled and bipedal humanoids using its sizable customer base. Its flagship is the PUDU D9, a full-sized biped built to serve in places like warehouses, stores, and hotels. It sells for between $20,000 and $30,000. A newer firm, AI Squared, recently closed a funding round worth hundreds of millions of yuan to rush its Alpha Bot 2 general-purpose humanoid to market. The startup, founded in 2023, says its AI combines slow reasoning and fast motion planning so its robots can plan, talk, and move simultaneously. An even newer firm, Lumos Robotics, has reportedly raised nearly $28 million in angel funding to develop a full-stack humanoid robotics platform. The second-generation iteration of its flagship robot, the LUS 2, recently demonstrated the ability to get up from the ground in just a second. According to Lumos, the next-gen bot features upgraded actuators and sensors for athletic-level dexterity. LimX Dynamics, known for its agile quadrupeds and bipeds, just shared footage showcasing its CL-3 humanoid’s lifelike walking gait and gestures. LimX says its flagship humanoid can watch human movement videos and replicate them via motion conversion, with no hand coding required. Just a year after its establishment, the startup DIGIT has unveiled multiple lines of sci-fi-inspired humanoids. Its Starwalker robots switch between legged and wheeled locomotion to cover more than 10,000 square meters across multiple floors. They’re available in five different colors. DIGIT also has a robot named Xia Lan with a humanlike face that can show a range of expressions and emotions. The leading Chinese robotics firm UBTECH also recently introduced its first hyper-realistic humanoid robot, Una, intended for applications like customer service and event marketing. UBTECH says it’s begun mass-producing its Walker S humanoid robots that are reportedly helping assemble iPhones for Foxconn. Footage of a so-called intelligent swarm of these humanoids at a Zeekr factory shocked the world. Another EV maker, Xpeng, plans to start mass-producing its Iron humanoids in 2026. The robot reportedly stole the show during the 2025 Shanghai Auto Week. The Shenzhen firm that’s gotten the most attention outside the usual tech echo chambers is Engine AI, which is planning the world’s first full-sized humanoid robot boxing tournament.

Mike Kalil

11,408 просмотров • 1 год назад

New framework: Kick down your robot, it will get back up every time 🥋 Chinese startup RoboParty is a Beijing startup founded April 2025 by Huang Yi, originally shipping ROBOTO ORIGIN, the world's first full-stack open-source bipedal humanoid. They released UFO: Unsupervised Reinforcement Learning Framework for Humanoid Control. DEFINITIONS -> what differs is where the learning signal comes from: - SUPERVISED: humans supply the right answers (labels), the model imitates them. - UNSUPERVISED: no answer key, the model finds structure in raw data on its own. - REINFORCEMENT LEARNING: no answer key either, the model tries things and a reward scores each attempt. → UNSUPERVISED RL: trial and error where the agent invents its own rewards, instead of engineers hand-writing one per task. REPRESENTATION LEARNING: compress raw states into a useful internal map. TEMPORAL DISTANCE: distance on that map is "how many steps from A to B." CONTRASTIVE: trained by pulling together what's close in time, pushing apart what isn't. -> CONTRASTIVE TEMPORAL-DISTANCE REPRESENTATION LEARNING: the model builds an internal map of body states where distance means how many steps it takes to get from one to another. It is trained by contrast: states that occur close together in a movement get pulled together in the map, randomly paired states get pushed apart. UFO is an open-source training framework that teaches humanoid robots skills, like getting up, walking, goal-reaching, teleoperation, without reference motions -> no motion-capture or human-video demonstrations to imitate. Its core is TeCH, a contrastive temporal-distance representation-learning algorithm: the robot explores, builds pseudo-goals by temporal rolling, and learns goal-conditioned policies from a single unified progress reward. One framework trains five different robots (Unitree G1/H1, RoboParty RP0/RP1, AgiBot X2) with automatic config conversion in ~2–3 hours per robot! The real novelty here "no demonstrations at all". No data-collection arms race,the dominant humanoid-locomotion recipe is tracking: imitate mocap/retargeted-human reference trajectories. The robot self-generates goals from its own exploration and learns from a progress reward, needing zero reference motion data. Everybody else is fighting over data acquisition, while this team just teleports out of the race entirely (inb4 "competition is for losers 💀 ). This strategy reminds me of the DeepSeek playbook applied to robots: open-source the whole stack to become the global default and commoditize everyone else. RoboParty is giving away hardware and now control software (UFO) to be the Android of humanoids. Yet another reason for the US to ban Chinese open models perhaps 🥶 ? What I also really like about this approach is the cross-embodiment infrastructure, one framework trains Unitree G1/H1, RoboParty RP0/RP1, and AgiBot X2 with automatic configuration conversion. Just like Physical Intelligence, RoboParty seems to place itself as a neutral hardware agnostic middle man. Also woth mentioning: their ability ot perform stable skill injection, e.g. adding a cartwheel without forgetting how to walk. A common failure of RL humanoid policies is that teaching a new agile skill destabilizes the existing ones (catastrophic forgetting). UFO claims you can inject rare motions (cartwheel) without collapsing learned behavior. If it holds, that's a significant incremental/continual skill-learning! But again, I have to underline it: no arXiv, no external validation, no success-rate numbers. -> robotics badely needs an independent unbiased evaluator imho. Still, look at that cool demo: robot is getting kicked and pushed around (serious disturbance) during teleoperation (controlled the person at the back wearing the VR headset), and still managed to always get back up. This is some serious demonstration of stability and robustness!

Léo

33,917 просмотров • 3 дней назад

Video: World’s first humanoid robot labor that swaps its own batteries to work endlessly | Jijo Malayil, Interesting Engineering Walker S2 uses dual-battery balancing and standardized modules to boost efficiency and ensure uninterrupted, optimized performance. In a leap for robotics, China’s UBTech has unveiled the Walker S2, the world’s first humanoid robot capable of fully autonomous battery swapping. Designed for non-stop industrial operations, the Walker S2 can replace its own power pack in just three minutes—no human intervention required. Equipped with advanced anthropomorphic bipedal locomotion and a hot-swappable battery system, Walker S2 is built to operate 24/7 across dynamic industrial environments. According to UBTech, the next-generation humanoid robot marks a major milestone in automation, bringing continuous, hands-free performance to the factory floor. In May 2025, UBTech Robotics and Huawei Technologies inked a significant partnership to accelerate the adoption of humanoid robots across China’s factories and households. Uninterrupted robot operations A video posted by the robotics firm opens with the sleek UBTech Walker S2 humanoid robot working in an industrial setting. The highlight, however, is its autonomous battery swap. Walker S2 approaches the charging station, carefully detaches its depleted power pack, and seamlessly installs a fresh one—all within about three minutes—without any human assistance, according to CGTN. The camera captures close-ups of the robot’s articulated limbs and the intelligent battery-handling mechanism, conveying precision and reliability. As the swap completes, Walker S2 resumes its duties, reinforcing the promise of uninterrupted, 24/7 operations in dynamic factory environments. UBTech’s Walker S2 humanoid robot is equipped with advanced dual-battery power balancing technology and uses standardized battery modules to optimize performance, reports CNEVPOST. This dual-battery system allows the robot to automatically switch to a backup battery in case of a main battery failure, ensuring that critical tasks are carried out without interruption. In addition to battery swapping, the robot can intelligently choose between charging and swapping based on task urgency, allowing it to manage energy dynamically and adapt to real-time operational demands. UBTech highlights these features as a step forward in deploying humanoid robots for industrial and domestic applications, combining flexibility, reliability, and autonomy in one intelligent platform. Factory intelligence upgrade Earlier in the year, UBTech unveiled a major advancement in humanoid robot collaboration, claiming the world’s first deployment of multiple humanoids working together across varied industrial tasks. Demonstrated at Zeekr’s 5G-enabled smart factory, the breakthrough centers on UBTech’s “BrainNet” framework, which orchestrates cooperative behavior through a cloud-device intelligence system. BrainNet integrates a “super brain” for high-level decision-making with an “intelligent sub-brain” for distributed multi-robot control. The super brain, powered by a proprietary large-scale multimodal reasoning model, handles complex production-line scheduling and decision-making. Meanwhile, the sub-brain coordinates real-time tasks using cross-field perception and Transformer-based control for dynamic adaptability. Together, they enable the Walker S1 humanoid robots to move beyond isolated operations and perform coordinated tasks with high precision and speed. The system is built on DeepSeek-R1 reasoning technology and trained on real-world data from automotive factory settings. Leveraging Retrieval-Augmented Generation (RAG), the model adapts to specific job functions and improves scalability across workstations. At Zeekr’s facility, dozens of Walker S1s now collaborate on tasks like assembly, inspection, and part handling. Using semantic VSLAM and shared mapping, they coordinate seamlessly via vision-based navigation and agile manipulation. UBTech says this marks a transition to “Practical Training 2.0,” where humanoid robots operate as a swarm, maximizing efficiency and setting the stage for next-generation intelligent manufacturing.

Owen Gregorian

35,637 просмотров • 1 год назад

China unveils humanoid robot worker with brain that runs 275 trillion ops/sec | Jijo Malayil, Interesting Engineering In tests, SUYUAN used vision and joint control to sort and move crates of various sizes, greatly improving warehouse productivity. Chinese manufacturing firm Shanghai Electric has unveiled its first self-developed industrial humanoid robot, “SUYUAN,” marking a major milestone in its robotics journey. Debuting at the World Artificial Intelligence Conference (WAIC 2025) on July 26 in Shanghai, SUYUAN boasts 38 degrees of freedom and 275 TOPS of on-device computing power, enabling precise operations and fluid movements. According to the firm, designed for diverse industrial use, the robot showcases Shanghai Electric’s end-to-end capabilities—from core tech to integrated solutions—and reinforces its commitment to next-gen industrial automation through a full industry chain strategy. At WAIC 2025, Shanghai Electric also unveiled a new joint venture with Johnson Electric for next-gen humanoid robotics and showcased its “LINGKE” dual-arm robot. Recently, Hangzhou-based Unitree Robotics launched the R1 humanoid with 26 joints for $5,900, showcasing athletic feats like cartwheels, running, and quick recovery. Smart factory assistant Shanghai Electric claims SUYUAN, equipped with 38 degrees of freedom (DoF) and a powerful 275 TOPS on-device computing processor, delivers fluid, human-like movements and high-precision operations across various industrial scenarios. Its advanced articulation and real-time processing capabilities make it highly adaptable, enabling smooth execution of complex tasks in dynamic work environments. SUYUAN, who weighs 110 pounds (50 kilograms) and is 5 feet 6 inches (167 cm) tall, was designed to have human-like proportions. Its 38-DoF articulation offers dexterity, allowing for both wide-range motion and sensitive manipulation. With a single arm, the robot can lift objects up to 4.4 pounds (2 kilograms) in weight and carry a total payload of up to 22 pounds (10 kilograms). With a walking pace of 3.1 miles per hour (5 km/h), SUYUAN is ideal for environments including assembly lines, warehousing, and logistics, according to a statement. To navigate complex industrial settings, SUYUAN combines LiDAR and binocular vision for self-guided mobility. Its 275-TOPS AI processor enables rapid data analysis and integration with large language models, allowing it to understand tasks in natural language and handle objects adaptively, reports Fox 44 News. In pilot demonstrations, the robot successfully identified, picked, and relocated crates of varying sizes using advanced computer vision and coordinated joint control—delivering measurable gains in warehouse efficiency. The company claims that SUYUAN’s launch represents a major turning point in Shanghai Electric’s foray into humanoid robotics and strengthens its vertically integrated approach to industrial automation solutions. Intelligent task handling Shanghai Electric also demonstrated its most recent developments in intelligent manufacturing at WAIC 2025, introducing a new joint venture with Johnson Electric centered on next-generation humanoid robotics and showcasing the “LINGKE” dual-arm robot. With its high-precision operations, adaptive teamwork, and closed-loop data capabilities, the LINGKE robot demonstrated live talents in handling complicated production jobs. LINGKE is made to do more than just replace human labor; it uses compliant force control and bimanual coordination to relieve workers of high-intensity, repetitive jobs. According to the company, the robot enhances operational efficiency by up to five times. Its core strength lies in a Data-Model-Deployment closed-loop system that starts with operational data, followed by data cleansing, model training, live deployment, and feedback-driven optimization—enabling autonomous learning and workflow improvement. Also at the event, Shanghai Electric and Johnson Electric introduced advanced hardware modules for humanoid robots, including rotary joints, linear joints, and dexterous finger joints. These components are designed to support smooth, precise, and quiet motion performance across robotics systems, reports Stock Titan. The joint venture announced two strategic agreements: a first-unit supply deal with the National and Local Co-Built Humanoid Robotics Innovation Center (Qinglong Project) and a cooperation memorandum with Fourier Robotics. Read more:

Owen Gregorian

51,638 просмотров • 1 год назад

Today we’re launching the first and only human-like AI agents in the world. Super Agents™ are the first agents with human‑level skills – they DM you, take @ mentions, send emails, manage docs, tasks, and more. Not just tools or API calls, but real skills fine‑tuned for how teams actually work. The first agents with 100% context – fully native in ClickUp and fully synced from other apps. Super Agents see your work the same way that humans do: tasks, docs, schedules, and conversations all in one place. The first agents that learn from human interactions automatically, without any setup or configuration – when you give feedback, they listen and improve how they work. The first agents with human‑level memory for custom agents – historical memory for every interaction, short-term working memory, and even long‑term memory stored in docs you can literally open, inspect, and edit. The first agents that are literally the same as users – our agentic user model is the same as our user data model. This gives you permissions and capabilities that you and your systems are already familiar with. The first infinite agent catalog – where anyone can create and customize agents in minutes, for literally any type of work imaginable. It's the most intuitive way to build agents on the planet. 95% of companies are failing in AI adoption. The reality is that AI isn't meant to be adopted, it's meant to be adapted – to you. Super Agents are automatically personalized to you and your company using proprietary state-of-the-art agent architecture, orchestration, and tooling. Today is the largest step forward we've ever made towards our mission of making people more productive. Maximize human productivity, with ClickUp Super Agents. Available NOW. For everyone.

Zeb Evans

320,607 просмотров • 7 месяцев назад

We are excited to share our latest work, "Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning," done in collaboration with Google DeepMind . Autonomous drones have reached superhuman speed in isolation, but what happens when multiple agents share the same airspace? Paper: Website: Video: Using league-based self-play, we train #ReinforcementLearning agents that race against a diverse, evolving population of opponents. Through this competitive training, sophisticated behaviors emerge without explicit programming: strategic overtaking, proactive collision avoidance, and even awareness of aerodynamic downwash from nearby drones. In real-world multi-player races at speeds exceeding 80kph (50 mph) and accelerations up to 7g, our agents outperform a five-time Swiss national drone racing champion while reducing collision rates by 50% compared to single-agent baselines. Crucially, training against diverse artificial opponents enables zero-shot generalization to human pilots, achieving over 90% race completion in mixed human-AI races with up to four competitors. A key insight: human pilots adopt riskier strategies when trailing, leading to more crashes under competitive pressure. Our learned policies, by contrast, maintain consistent safety margins regardless of race standing, a property essential for deploying autonomous systems alongside humans. Also, the multi-agent self-play policies are more robust than those trained independently, suggesting that training in competitive environments is not only key to winning races but also to learning safer, more reliable autonomy for real-world multi-robot systems. Kudos to Ismail Geles, Leonard Bauersfeld, Markus Wulfmeier! Ismail Geles Leonard Bauersfeld Markus Wulfmeier European Research Council (ERC) UZH IfI University of Zurich UZH Science UZH Space Hub Swiss Robotics NCCR Robotics

Davide Scaramuzza

14,612 просмотров • 2 месяцев назад