Загрузка видео...

Не удалось загрузить видео

На главную

Generalist AI has just broken the "impossible triangle" of speed, reliability, and intelligence in robotics. They’ve pushed their embodied foundation model forward in a big way: GEN-1 hits “Mastery” on simple physical tasks,99% average success rate, and it can improvise when things go wrong. Smooth, natural, fast(factory owners must...

25,519 просмотров • 4 месяцев назад •via X (Twitter)

Комментарии: 0

Нет доступных комментариев

Здесь появятся комментарии из оригинального поста

Похожие видео

My conversation with Sergey Levine (Sergey Levine). Sergey is the co-founder of Physical Intelligence -- a company building foundation models that can control any robot to do any task in any environment. The company's thesis is that generality is more scalable than specialization, meaning that a model trained across many different robots and tasks will ultimately outperform any system built to do one thing well (eg, just wash dishes). Sergey is a researcher by background, but I think you will appreciate how practical and commercially grounded this conversation is. We discuss: - Why changing a diaper will be the last task a robot masters - The simulation v. real-world data debate - How multimodal LLMs give robots common sense - Moravec's Paradox + Robot Olympics - Why robots can do long-horizon tasks now - A realistic timeline for robots in our homes I should note that I am an investor in Physical Intelligence -- I made the investment because I believe it is one of the most important companies tackling the problem of robotics. Enjoy! Timestamps: 0:00 Intro 2:39 Defining Physical Intelligence 5:19 The Challenge of Building General Models 6:34 The Stakes and Future of General Purpose Robotics 8:15 Pros and Cons of Humanoid Robots 10:12 Historical Milestones in Robotics Research 15:31 Combining Generative AI and Deep RL 21:24 Moravec's Paradox 25:33 Kitchen Robots 29:30 Simulation vs. Real-World Data 30:48 The Robot Olympics 36:31 The Physiological Reality of Embodiment 38:56 Controversies in the Robotics Community 44:18 What Makes a Great Researcher 48:27 How Businesses Should Prepare for Robotics 54:09 Tracking Progress Through Research Papers 57:02 The Next Step: Mid-Level Reasoning 1:02:00 The Kindest Thing

Patrick OShaughnessy

133,833 просмотров • 4 месяцев назад

Video: World’s first humanoid robot labor that swaps its own batteries to work endlessly | Jijo Malayil, Interesting Engineering Walker S2 uses dual-battery balancing and standardized modules to boost efficiency and ensure uninterrupted, optimized performance. In a leap for robotics, China’s UBTech has unveiled the Walker S2, the world’s first humanoid robot capable of fully autonomous battery swapping. Designed for non-stop industrial operations, the Walker S2 can replace its own power pack in just three minutes—no human intervention required. Equipped with advanced anthropomorphic bipedal locomotion and a hot-swappable battery system, Walker S2 is built to operate 24/7 across dynamic industrial environments. According to UBTech, the next-generation humanoid robot marks a major milestone in automation, bringing continuous, hands-free performance to the factory floor. In May 2025, UBTech Robotics and Huawei Technologies inked a significant partnership to accelerate the adoption of humanoid robots across China’s factories and households. Uninterrupted robot operations A video posted by the robotics firm opens with the sleek UBTech Walker S2 humanoid robot working in an industrial setting. The highlight, however, is its autonomous battery swap. Walker S2 approaches the charging station, carefully detaches its depleted power pack, and seamlessly installs a fresh one—all within about three minutes—without any human assistance, according to CGTN. The camera captures close-ups of the robot’s articulated limbs and the intelligent battery-handling mechanism, conveying precision and reliability. As the swap completes, Walker S2 resumes its duties, reinforcing the promise of uninterrupted, 24/7 operations in dynamic factory environments. UBTech’s Walker S2 humanoid robot is equipped with advanced dual-battery power balancing technology and uses standardized battery modules to optimize performance, reports CNEVPOST. This dual-battery system allows the robot to automatically switch to a backup battery in case of a main battery failure, ensuring that critical tasks are carried out without interruption. In addition to battery swapping, the robot can intelligently choose between charging and swapping based on task urgency, allowing it to manage energy dynamically and adapt to real-time operational demands. UBTech highlights these features as a step forward in deploying humanoid robots for industrial and domestic applications, combining flexibility, reliability, and autonomy in one intelligent platform. Factory intelligence upgrade Earlier in the year, UBTech unveiled a major advancement in humanoid robot collaboration, claiming the world’s first deployment of multiple humanoids working together across varied industrial tasks. Demonstrated at Zeekr’s 5G-enabled smart factory, the breakthrough centers on UBTech’s “BrainNet” framework, which orchestrates cooperative behavior through a cloud-device intelligence system. BrainNet integrates a “super brain” for high-level decision-making with an “intelligent sub-brain” for distributed multi-robot control. The super brain, powered by a proprietary large-scale multimodal reasoning model, handles complex production-line scheduling and decision-making. Meanwhile, the sub-brain coordinates real-time tasks using cross-field perception and Transformer-based control for dynamic adaptability. Together, they enable the Walker S1 humanoid robots to move beyond isolated operations and perform coordinated tasks with high precision and speed. The system is built on DeepSeek-R1 reasoning technology and trained on real-world data from automotive factory settings. Leveraging Retrieval-Augmented Generation (RAG), the model adapts to specific job functions and improves scalability across workstations. At Zeekr’s facility, dozens of Walker S1s now collaborate on tasks like assembly, inspection, and part handling. Using semantic VSLAM and shared mapping, they coordinate seamlessly via vision-based navigation and agile manipulation. UBTech says this marks a transition to “Practical Training 2.0,” where humanoid robots operate as a swarm, maximizing efficiency and setting the stage for next-generation intelligent manufacturing.

Owen Gregorian

35,637 просмотров • 1 год назад

Karol Hausman is the co-founder and CEO of Physical Intelligence, a robotics company building a general-purpose “AI brain for the physical world.” The company has raised more than $1 billion in funding to develop foundation models that allow robots to operate across many machines, environments, and tasks rather than being programmed for a single purpose. In our conversation, we explore: • The moment a lecture from Sergey Levine convinced him to abandon his PhD research direction and pivot fully to deep learning • The case for building a general “AI brain” for the physical world rather than a single specialized robot • The role of real-world data in training robots, the limits of simulation, and how deployment could create a powerful data flywheel • The unique challenges of physical intelligence and why robots must operate with far higher reliability than language models Thank you to the partners who make this possible - Brex: The intelligent finance platform: - Granola: The app that might actually make you love meetings: Timestamps (00:00) Intro (04:05) Karol’s early fascination with robots (18:21) Karol’s entry point to robotics and PhD program (25:49) Combining robotics with LLMs: The Taylor Swift demo (30:48) The 1970s SHRDLU AI experiment (39:40) How research shapes what Physical Intelligence builds (49:07) The return of reinforcement learning in robotics (1:00:00) NVIDIA’s simulation engines (1:07:31) Compensating for missing senses

Mario Gabriele 🦊

27,871 просмотров • 4 месяцев назад

Excited to announce GR00T N1, the world’s first open foundation model for humanoid robots! We are on a mission to democratize Physical AI. The power of general robot brain, in the palm of your hand - with only 2B parameters, N1 learns from the most diverse physical action dataset ever compiled and punches above its weight: - Real humanoid teleoperation data. - Large-scale simulation data: we are open-sourcing 300K+ trajectories! - Neural trajectories: we apply SOTA video generation models to “hallucinate” new synthetic data that features accurate physics in pixels. Using Jensen’s words, “systematically infinite data”! - Latent actions: we develop novel algorithms to extract action tokens from in-the-wild human videos and neural generated videos. GR00T N1 is a single end-to-end neural net, from photons to actions: - Vision-Language Model (System 2) that interprets the physical world through vision and language instructions, enabling robots to reason about their environment and instructions, and plan the right actions. - Diffusion Transformer (System 1) that “renders” smooth and precise motor actions at 120 Hz, executing the latent plan made by System 2. We deploy N1 on GR1 robot, 1X Neo robot, and a large collection of simulation benchmarks. N1 achieves up to +30% boost in diverse manipulation tasks for household and industrial settings. While humanoid robots are the main focus of N1, our model also supports cross-embodiment. We finetune it to work on the $110 HuggingFace LeRobot SO100 robot arm! Open robot brain runs on open hardware. Sounds just right. Let’s solve robotics, together, one token at a time. Links to our Whitepaper, Github repo, HuggingFace model, and open dataset page in the thread: 🧵

Jim Fan

466,333 просмотров • 1 год назад

A Few Thoughts on Robotics The criticism that robotics can only be used in a rather one-sided way is, at the same time, the solution to the problem. What do I mean by that? Since the Industrial Revolution, humanity has increasingly made production methods more efficient. Fordism introduced assembly line work, but this comes at the expense of monotonous, repetitive tasks. On the one hand, immense wealth has been created; on the other hand, countless people suffer from repetitive tasks, which are a direct consequence of that industrial revolution and the division of labor- in other words, assembly line work. The debate about whether AI and robotics could impact the labor market is answered in different ways. I have a clear opinion on this: Up to now, technology has merely been an augmentation, an improvement of human labor to make it more effective. Robotics and AI, however, represent a qualitative break with this situation. For the first time in human history, it won't be humans who become more efficient, but rather replaceable, insofar as human augmentation becomes *less* efficient than replacing human labor with robotics. In just a few years, a human using technology will simply be less efficient than a robot that doesn't know an eight-hour day, weekends, or holidays, but can perform monotonous tasks 24/7 on an assembly line without breaking down due to physical ailments or needing medical attention. Wear and tear simply means replacing specific parts of the robot. To return to the initial question: production doesn't require general-purpose robots capable of performing a wide variety of tasks, but rather specialized robots that excel at the specific tasks for which they are needed. Figure02 vividly illustrates why this is only now possible: even the simplest assembly line work still requires delicate manual dexterity because the production line is designed for human hands. This breakthrough has now arrived, but AGI (Automated Generating Intelligence) isn't necessary for robots to be used in production processes. It's sufficient that they can perform monotonous tasks. And that's why I believe 2026 will be the year of the robots. (Clip: Figure02 in production chain at BMW Car-production)

Chubby♨️

15,228 просмотров • 8 месяцев назад

A policy that teaches robot hands to touch things the way humans do... not just grab and move, but feel and adjust in real time. Robot manipulation research often stops at picking up objects and placing them. CGP goes further: it handles tasks like opening jars, flipping objects in-hand, wiping dishes, and grasping fragile eggs, the kind of dexterous, contact-rich skills that require constant micro-adjustments based on what the fingers are actually feeling. The robot doesn't just see what it's doing; it predicts what contact should feel like at each step, then checks whether reality matches the prediction. If a finger is slipping, the policy knows before the object drops. Works on real robot hands (both 4-finger and 5-finger designs) with tactile sensors embedded in the fingertips Robust to visual distractions! The robot keeps flipping a box correctly even when the camera view is disrupted, because it's grounding decisions in touch, not just vision. Baseline policies without contact grounding fail in predictable ways: slipping mid-task, incomplete motions, loss of grasp, CGP avoids these This is a meaningful step toward robots that can handle the physical world with the kind of reliable, adaptive grip that humans take for granted. Relevant for manufacturing, logistics, assistive robotics, and anywhere fragile or irregular objects need to be handled carefully. Published at RSS 2026, developed with Meta Reality Labs Research. Thanks for sharing, Zhengtong Xu / Zhengtong Xu ——- Weekly robotics and AI insights. Subscribe free:

Ilir Aliu

12,769 просмотров • 2 месяцев назад

Japan Just Built a HouseBot You Control Without Speaking and It Changes Everything! Donut Robotics has officially unveiled its first bipedal humanoid, Cinnamon 1, and instead of focusing on louder voices or bigger motors, the company went in the opposite direction. Silence. Cinnamon 1 introduces what Donut Robotics calls Silent Gesture Control, a system that allows the humanoid to be guided using simple hand and finger movements rather than spoken commands. This approach feels especially well suited for real world environments where traditional voice control falls apart. Busy factory floors. Construction sites filled with constant noise. Even quiet indoor settings where voice commands feel awkward or intrusive. It also opens the door for far more accessible human robot interaction, particularly for users with impairments. While the current Cinnamon 1 hardware is built on an OEM platform, the intelligence driving it is where Donut Robotics is placing its long term bet. The team is actively developing custom Vision Language Action AI that allows the robot to interpret what it sees, understand intent, and respond with physical action. The goal is not just smarter robots, but robots that feel more natural. Even more ambitious is the company’s plan for full domestic production. Donut Robotics has stated its intention to localize both manufacturing and AI development in Japan, reinforcing the country’s reputation for precision engineering and thoughtful robotics design. If timelines hold, Cinnamon 1 units are expected to begin deployment in factories and construction environments by the end of 2026. That puts this humanoid squarely in the category of near term reality rather than distant concept. The takeaway is simple but important. As humanoid robots move out of labs and into daily work environments, the winners may not be the loudest or flashiest machines. They may be the ones that understand us without a word being spoken.

The AI Robot Guy on X

257,928 просмотров • 6 месяцев назад

AheadForm just raised a new A1 round worth hundreds of millions of RMB (~tens of millions USD) 🤖 That matters because this is not another humanoid company chasing locomotion first. AheadForm is building around the part most robotics startups still underestimate: face, emotion, and real-time human connection. The new funding will go into multimodal embodied interaction, emotion foundation models, facial hardware and materials, standardized delivery, and global expansion. Founded in June 2024, the company is still young, but the founder’s research trail is not. Yuhang Hu, a Columbia PhD and AheadForm’s CEO/CTO, has published work spanning facial coexpression, realistic lip motion for humanoid face robots, and self-supervised robot self-modeling. That is the deeper signal here. In a market crowded with hands, arms, and walking demos, investors are now putting serious money behind embodied AI that can express, respond, and hold attention face to face. And the company is moving fast. According to public reports, AheadForm has completed five funding rounds since the second half of 2025, while its robots have already broken out of lab-only visibility through public activations like the NetEase Justice mobile game collaboration and large robot-stage appearances. If humanoid robotics is about physical labor, AheadForm is making the case that the next layer is emotional presence. That may end up being one of the more important categories in embodied AI.

RoboHub🤖

155,548 просмотров • 4 месяцев назад