Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Gemini Robotics ER 2 is our most capable embodied reasoning model designed for physical AI 🤖 Built as a high-level brain for robotics, the model connects directly to the Gemini Live API. It processes continuous video streams to track progress, call tools, search the web, and command an action...

17,581 görüntüleme • 2 gün önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

Most robotics AI models suffer from the "stop-and-think" problem. They take a static picture, pause to reason, execute an action, and repeat. In the real world, that latency causes spills, collisions, and failed tasks. Google DeepMind just launched Gemini Robotics ER 2: an embodied reasoning model that thinks and acts at the speed of the physical world. Here's why this is a step-change for physical AI engineering: Traditional robotics models rely on static snapshots. But knowing *when* a task is done, such as when to stop pouring coffee into a cup or when a trash bag is securely tied, requires continuous temporal awareness. Gemini Robotics ER 2 integrates directly with the bidirectional streaming Gemini Live API to reason about what comes next while simultaneously executing motor actions. What makes Gemini Robotics ER 2 different: 🎯 91.3% accuracy on live video moment-finding (0.96s mean absolute distance) at 4x the execution speed of frontier models 📈 Continuous progress tracking across 5 completion stages (57.4% accuracy) to self-correct mid-task without restarting 🛠️ Native agentic tool orchestration that commands lower-level VLA models, navigation APIs, and Google Search 🤝 Multi-robot collaboration allowing physically diverse machines (like Apptronik's Apollo 2 humanoid and Franka's FR3 Duo arm) to hand off tasks in shared spaces 🛡️ Built-in physical safety that autonomously halts robots when humans enter a workspace and resumes once clear

Karl Weinmeister

26,179 görüntüleme • 2 gün önce