Loading video...

Video Failed to Load

Go Home

Without World Models, There Is No AGI. Google Just Proved It. If AGI ever happens, it will not come from bigger chatbots alone. From the very start of this interview, one thing is crystal clear: without world models, we will never reach AGI. And right now, Google is leading...

23,784 views • 7 months ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

Demis Hassabis on the limit in today’s AI: language can describe the world, but it cannot contain it - and why "World Models" are his "longest standing passion". Language models absorbed far more structure about reality from text than many researchers expected, because human language quietly carries physics, psychology, culture, tools, plans, and cause-and-effect. But text is still a compressed residue of experience, not experience itself. A sentence can say a cup falls from a table, yet it does not fully encode weight, grip, balance, friction, timing, sound, surprise, or the tiny motor corrections a body makes before it even notices them. The world is not only made of facts that can be named; it is made of constraints that have to be lived through, touched, predicted, violated, and repaired. That is why world models matter. They aim to learn the hidden grammar of physical reality: how objects persist, how forces unfold, how space changes when an agent moves, and how action creates feedback. Language models can often reason about the world because people have written so much about it. World models try to learn what the world is like before it becomes words. The difference is exactly what matters because intelligence is not just answering well; it is knowing what would happen next if you moved, reached, pushed, smelled, slipped, or failed. A mind trained only on descriptions may become brilliant at explanation. A mind trained on experience may become better at consequence. --- Full video from "Google DeepMind" and "Hannah Fry" YT channel (link in comment)

Rohan Paul

49,938 views • 2 months ago

The interview with Demis Hassabis - the tl;dr (summary) about scaling, AGI and much more: 1. Solving the "Root Node" Problems: DeepMind isn't just building chatbots; they are using AI to solve the hardest scientific problems. After the success of AlphaFold, they are now targeting materials science (room-temperature superconductors, better batteries) and even nuclear fusion to unlock unlimited clean energy. 2. The "Jagged Intelligence" Paradox: Current AI models are in a weird spot—they can win gold medals at the International Math Olympiad but still fail at basic logic puzzles. Hassabis calls this "jagged intelligence." The goal isn't just more data, but fixing these inconsistencies to make models reliable across the board. 3. Scaling is Not Dead (But it’s Changing): Despite rumors of hitting a "data wall," Hassabis says we haven't seen a hard limit yet. However, we are seeing diminishing returns. His bet? Getting to AGI will require 50% scaling and 50% architectural innovation. It’s no longer just about making the models bigger; it’s about making them smarter. 4. The Missing Piece: System 2 Thinking: Today's models are passive—they just spit out an answer. To reach AGI, we need systems that can "think" before they speak. This involves planning, reasoning, and double-checking their own work (similar to human "System 2" thinking) rather than just predicting the next word. 5. Rise of World Models: The next big frontier is "World Models" (like their project Genie). AI needs to understand the physics of the world—gravity, object permanence, and cause-and-effect—not just language. This is crucial for building helpful digital agents and robots that can navigate real-life situations. 6. Is the Universe Computable? On a philosophical level, Hassabis believes that everything in the universe might be computable. His life's work is testing the limits of the "Turing Machine." If we can build an AGI that simulates the human mind perfectly, we might finally understand what (if anything) makes human consciousness unique. 7. Bigger than the Industrial Revolution: We need to prepare for a shift that is 10x faster and bigger than the Industrial Revolution. If AI solves energy (fusion) and labor, we might enter a "post-scarcity" world. Hassabis warns that society, economics, and governments need to adapt quickly to ensure these benefits are shared by everyone, not just a few. And since this is the most important aspect, here is the clip about post labor economy:

Chubby♨️

27,473 views • 7 months ago

This is THE moment of Physical AI! We are officially announcing Cosmos 3: Omnimodal World Models for Physical AI 🚀 - Cosmos 3 is an omnimodal world model: within a unified architecture, it can understand and generate language, images, video, audio, and actions. - It is not just a VLM, not just a video generator, not just an audio-visual generative model, and not just a physics simulator / world-action model. It can understand images and videos, generate images, videos, and audio, simulate future worlds, predict actions, and generate robot policies—enabling models to truly begin to “touch the world.” - Cosmos 3 is the #1 open-weight reasoner / T2I / I2V / robot policy across many benchmarks. Huge thanks to every teammate who fought side by side on this journey—from architecture, data, training, infra, serving, and evaluation to post-training. Every part of this project carries an incredible amount of hard work. This was my first time leading a project as Tech Lead, and I feel truly fortunate. The future of Physical AI needs models that can not only “see” and “describe” the world, but also “imagine,” “simulate,” and “act”—and eventually close the loop with the real world. I hope Cosmos 3 can become an important starting point for this direction, and I’m excited to push Physical AI into its next stage together with the open-source community. Welcome to the era of Physical AI. HuggingFace: Project Website: Code:

Max Zhaoshuo Li 李赵硕

1,078,268 views • 2 months ago