Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

NEWS: Waymo has introduced the Waymo World Model, "a frontier generative model built on Google DeepMind’s Genie 3 that sets a new bar for large-scale, hyper-realistic autonomous driving simulation." "By simulating the “impossible”, we proactively prepare the Waymo Driver for some of the most rare and complex scenarios—from tornadoes...

133,382 Aufrufe • vor 7 Monaten •via X (Twitter)

34 Kommentare

Profilbild von Yun-Ta Tsai
Yun-Ta Tsaivor 7 Monaten

One problem of such world model is that you have to generate hosts of the sensor suite data which ideally minimizes train/test mismatch. It is harder to generate LiDAR than RGB in physical accuracy in terms of photometric response — and you are coupled with very specific band pass filter and emitter to a very specific brand while considering the real material reflectance. It might generate something looks reasonable, but introduces bias to the training. One could argue that it could also overcome with augmentation but modeling noise perturbation of photons had been genuinely hard due to the quantum mechanics. Any extra layer of sensor generation limits your scale.

Profilbild von Ascraeus.Tharsis
Ascraeus.Tharsisvor 7 Monaten

Maybe try simulating power outages...?

Profilbild von TeslaFUDKer 🍁⚡️ 𝕏𝕏𝕏
TeslaFUDKer 🍁⚡️ 𝕏𝕏𝕏vor 7 Monaten

So they are finally copying Tesla. Got it 😂

Profilbild von Scott T Archer
Scott T Archervor 7 Monaten

About time Waymo copies Tesla

Profilbild von robot*irl
robot*irlvor 7 Monaten

They might catch up to Tesla after all! Just 7 Billion miles to go🌞

Profilbild von JC Christopher
JC Christophervor 7 Monaten

An AI-based real world model. I hope Waymo keeps going this direction. We all know what's going to eventually happen as this gets better and better for them. It may take a different management team with less pride, but it will happen.

Profilbild von Dennis ♨
Dennis ♨vor 7 Monaten

Maybe they should start by simulating REALISTIC SCENARIOS first.. like telephone poles and having kids running out from SCHOOL BUSSES.

Profilbild von B.O.H.I.C.A
B.O.H.I.C.Avor 7 Monaten

Didn't Elon Musk JUST say that they already have that, like this week?

Profilbild von Phil Trubey
Phil Trubeyvor 7 Monaten

Waymo showed us their perception capabilities. Do they still have different stacks for producing driver inputs like steering and braking? Ie. My understanding is that they still do not have a fully end to end AI model for their cars?

Profilbild von Stewart Believer
Stewart Believervor 7 Monaten

Hate to be that guy, but let me know when they buy a company that actually builds cars at scale.

Profilbild von Brian
Brianvor 7 Monaten

Maybe they should simulate some more normal stuff, like fire trucks, construction zones and power outages.

Profilbild von Tech Equity Engineer
Tech Equity Engineervor 7 Monaten

Modern simulators aim to reduce bias in the training through careful design and real data integration. Whether or not folks agree with what Waymo is doing, they are doing millions of miles per week and continuing to scale and that is a fact.

Profilbild von beefcube (execute/commies)⚔️
beefcube (execute/commies)⚔️vor 7 Monaten

utterly useless and waste of money

Profilbild von JustGiveMeAGreatL2
JustGiveMeAGreatL2vor 7 Monaten

Wait, I thought you needed 10B miles? Maybe Waymo knows what they are doing

Profilbild von WangNextDoor
WangNextDoorvor 7 Monaten

Great news, but you're still 3 years behind Nvidia on real-world simulation and probably 6 years behind Tesla. Sorry but yeah, nice try!

Profilbild von Simon P
Simon Pvor 7 Monaten

Simulating rare edge cases is a smart move for safety. While real-world driving data is usually the best teacher, these models can help cover the 'impossible' scenarios. It is interesting to see how different companies are using AI to solve the autonomy puzzle.

Profilbild von Hugo Not The Boss
Hugo Not The Bossvor 7 Monaten

But the world model is based on vision only. So how the hech do they train their lidar???

Profilbild von EscapeTheGreatFilter
EscapeTheGreatFiltervor 7 Monaten

Strange how Waymo is releasing this right after Tesla showed similar tech off 😂 Everyone siloing everything until it's made public

Profilbild von Denis Ulmer
Denis Ulmervor 7 Monaten

@grok explain how this is different or better / worse than what Tesla is doing for years now.

Profilbild von Overly Trev’s Future
Overly Trev’s Futurevor 7 Monaten

They must of read Ashoks article lol

Profilbild von Mal
Malvor 7 Monaten

“Prepare the Waymo driver” oh you mean the random Philipino 5k miles away, got it

Profilbild von Derek | Living Autonomous
Derek | Living Autonomousvor 7 Monaten

Good for them and all. They will learn from Tesla and do fine for themselves.

Profilbild von Raj
Rajvor 7 Monaten

But Lidar

Profilbild von William Weber
William Webervor 7 Monaten

Waymo will last about 2-3 years before closing shop. They will lose many billions in the process.

Profilbild von Pedro Pallotta - Space Orbit
Pedro Pallotta - Space Orbitvor 7 Monaten

Philippines liked this post

Profilbild von Damon Wolfgang Gnojek
Damon Wolfgang Gnojekvor 7 Monaten

Remember when Tesla showed us their world model at like AI day 2 (maybe 1…), many years ago?

Profilbild von Jive Turkey
Jive Turkeyvor 7 Monaten

Better late than never.

Profilbild von Arpe
Arpevor 7 Monaten

So many companies have these. Why has nobody of them made it to a video game to take on GTA6 or Gran Turismo 🤔

Profilbild von Cornfed_88
Cornfed_88vor 7 Monaten

The problem is they cant envision every possible edge case using sim only. Sure it can help, but you need real world data to solve the problem fully.

Profilbild von Ryan Wang 🇹🇼
Ryan Wang 🇹🇼vor 7 Monaten

Waymo is using Gemini's World Model for training, and it's absolutely impressive! However, Tesla is doing the same thing too. A lot of people are underestimating it — Tesla actually has one of the best World Models out there.

Profilbild von James Stevens 🏴󠁧󠁢󠁥󠁮󠁧󠁿🇬🇧
James Stevens 🏴󠁧󠁢󠁥󠁮󠁧󠁿🇬🇧vor 7 Monaten

bet they wish they'd recorded driving data & video when they ran all the google earth photo cars, tho

Profilbild von Cyber Pigeon's Research
Cyber Pigeon's Researchvor 7 Monaten

Keep using lidar🤣

Profilbild von randomhamlife
randomhamlifevor 7 Monaten

Or… gather 8 billion ACTUAL miles of multi-camera video. No dice, Waymo. It takes more than video game programming to make a functioning, real-world AI.

Profilbild von Kevin Pham
Kevin Phamvor 7 Monaten

good hope waymo get better so the ev market expand faster

Ähnliche Videos

NEWS: Waymo has released a new blog post detailing their AI strategy and how it’s allowing them to bring service to more riders faster. "Achieving demonstrably safe AI — where safety is proven, not just promised — requires a holistic approach. Beyond a smart and capable Driver, you also need a closed-loop, realistic Simulator to train and rigorously test the Driver in a myriad of challenging situations, and a sharp Critic to evaluate the Driver's performance and identify areas for improvement." Waymo says that autonomous driving isn’t just a matter of building a “smart driver,” but rather creating a full AI ecosystem centered on safety from the ground up. At the core is the Waymo Foundation Model, a unified world-model that powers all major components of Waymo’s autonomous stack (Driver, Simulator, Critic). "By using a “Think Fast/Think Slow” architecture (combining rapid sensor-fusion with deep semantic reasoning), this system enables the car to detect complex and rare road scenarios (e.g. a burning vehicle ahead), reason about them, and choose safe behavior. Waymo trains large “Teacher” AI-models for driving, simulation, and evaluation, then distills them into smaller, efficient “Student” models suitable for real-world deployment, while keeping safety validation tightly integrated. The result is a continuous “flywheel” of learning: driving data (real and simulated) generate feedback, which leads to refinements, more simulation, more data, and only when safety checks pass is new code deployed. Having already exceeded 100 million fully autonomous miles, Waymo reports a more than ten-fold reduction in severe-injury crashes compared to human drivers." Full blog post:

Sawyer Merritt

83,635 Aufrufe • vor 9 Monaten

Today we're announcing #GAIA1: a 9B parameter world model, trained on 4,700 hours of driving data, able to simulate complex and diverse driving scenes from video, text and action inputs. This model is 480x larger than the preview we shared earlier this year and the results are incredible. These videos are entirely synthetically generated by Wayve's generative AI, GAIA-1. But there is more here than just generating videos, GAIA is an entire world model. A world model allows us to simulate the future, conditioned on video, text and action inputs, which can be leveraged for making informed decisions when driving. Why is this game-changing for autonomous driving? 1. Safety. One limitation with AI systems like today's Large Language Models is that they are autoregressive, next-word prediction algorithms, but aren't necessarily aware of the implications of their decisions. A world model allows us to give our AI the capability to be aware of its decisions, by simulating the future, which is important for self-driving safety. 2. Synthetic training data. I believe synthetic training data is the future for AI, because it is safer, cheaper, and infinitely scalable. GAIA-1 unlocks unprecedented realism and diversity of synthetic data for self-driving. 3. Long-tail robustness. One of the biggest challenges for self-driving is long-tail robustness: dealing with the enormous magnitude of edge cases we see on the road. An advantage of generative AI is its incredible ability to recombine experiences in new ways. This is exciting for self-driving as it means we can learn from two edge case scenarios, and combine them to become a corner case. For example, we can experience driving in fog, and experience of jay-walking pedestrians, and GAIA can learn from these experiences to understand how to generate a fog+jay walking scenario. Check out many more videos in our blog or further technical details in our paper: Or come chat with our team who are at the International Conference on Computer Vision (#ICCV2023) this week in Paris in Booth 32 Jamie Shotton

Alex Kendall

631,909 Aufrufe • vor 3 Jahren

Tencent presents GameGen-O Open-world Video Game Generation We introduce GameGen-O, the first diffusion transformer model tailored for the generation of open-world video games. This model facilitates high-quality, open-domain generation by simulating a wide array of game engine features, such as innovative characters, dynamic environments, complex actions, and diverse events. Additionally, it provides interactive controllability, thus allowing for the gameplay simulation. The development of GameGen-O involves a comprehensive data collection and processing effort from scratch. We collect and build the first Open-World Video Game Dataset (OGameData), amassed extensive data from over a hundred of next-generation open-world games, employing a proprietary data pipeline for efficient sorting, scoring, filtering, and decoupled captioning. This robust and extensive OGameData forms the foundation of our model's training process. GameGen-O undergoes a two-stage training process, consisting of foundation model pretraining and instruction tuning. In the first phase, the model is pre-trained on the OGameData via the text-to-video and video continuation, endowing GameGen-O with the capability for open-domain video game generation. In the second phase, the pre-trained model is frozen, and we fine-tuned using a trainable InstructNet, which enables the production of subsequent frames based on multimodal structural instructions. This whole training process imparts the model with the ability to generate and interactively control content. In summary, GameGen-O represents a notable initial step forward in the realm of open-world video game generation via generative models. It underscores the potential of generative models to serve as an alternative to rendering techniques, which can efficiently combine creative generation with interactive capabilities.

AK

367,249 Aufrufe • vor 2 Jahren

Applied Intuition CEO Qasar Younis on the autonomous driving approaches being taken by Tesla and Waymo, and why both can succeed in the years to come: "Generally speaking, every single car company on the planet right now is working on a product that’s like a Tesla FSD product. Many companies are working on versions of that, that would become fully autonomous within a cheap sensor suite." "So the fundamental difference, just to simplify the Tesla approach versus the Waymo approach... the Waymo approach is lots of sensors and lots of compute, and maps, and the Tesla version is very few sensors, no high fidelity maps... and cheaper compute for a lack of a better word." "And the Tesla version of a product, this in the industry is called an L2++ product, is going to be available everywhere because it’s literally cheaper and it doesn’t require HD maps. The Waymo product functions better in a geographically constrained area." "So, fast forward five years, both of these types of technologies will be much more ubiquitous. L2++ and L4 will be much more ubiquitous, not only in the Bay Area or in parts of China, but really globally - there are companies working on this globally." "We’re at that moment for L2++ systems. Where people are willing to pay thousands of dollars for a semi-automated vehicle." "It will not be a long time — you’re already seeing this in China — where the downward pricing pressure for the autonomous product, for lack of a better word, will become close to free." Qasar Younis Lenny Rachitsky

a16z

60,385 Aufrufe • vor 6 Monaten