Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Demis Hassabis just explained why the real AI bottleneck has nothing to do with training runs. Most people picture the AI arms race as who can build the biggest model. GPT-4 or Gemini Ultra style training runs, a few hundred million in compute, fired once or twice a year....

32,150 Aufrufe • vor 5 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

Jonathan Ross just revealed why AI companies aren’t growing faster. Not demand. Not competition. Physics. Ross: “The demand for compute is insatiable.” There isn’t enough compute in the world. Not a temporary shortage. A fundamental gap between what the market wants and what the infrastructure can deliver. Ross: “Right now, one of the biggest complaints of Anthropic is the rate limits. People can’t get enough tokens.” Rate limits aren’t product decisions. They’re rationing. Companies forced to regulate access because infrastructure cannot meet demand. Slower services. Token caps. The only things standing between these companies and a revenue surge they can’t access. Every token cap is a revenue cap. Every slowdown is a sale that didn’t happen. Ross: “If Anthropic was given twice the inference compute, within one month their revenue would almost double.” Read that again. Double the compute. Double the revenue. Within thirty days. That’s not a growth projection. That’s a measurement of how deep the backlog already is. The demand exists right now. It’s sitting in a queue. The only thing between these companies and that revenue is physical hardware they don’t have. This breaks every assumption about how tech companies scale. Usually you scale by finding customers. AI companies have infinite customers. They scale by finding hardware. The constraint isn’t market fit. It isn’t distribution. It isn’t competition. It’s processing power. This is why Jensen Huang is the most important person in the world right now. NVIDIA doesn’t just make chips. It makes the thing every government, every AI lab, and every company racing for this future needs more of and can’t get enough of. The compute bottleneck isn’t a tech industry problem. It’s a civilizational one. The winner of this era isn’t determined by who builds the smartest model. Every major lab has a frontier model. The winner is whoever secures the most compute fastest while everyone else rations what’s left. The race isn’t for intelligence. It’s for infrastructure. And right now there isn’t enough to go around.

Dustin

28,395 Aufrufe • vor 7 Monaten

The entire AI boom has hit a physical wall that no chip can break: America does NOT have enough electricity to run what these companies are building. You don't form an emergency coalition around a problem you've solved. You form one around a problem that's beating you. This morning Nvidia, Google, and a startup called Emerald AI launched something called the AI Energy Management Alliance. 18 members, including the AI lab Anthropic, the utility National Grid, and the power producers AES and NRG. The stated mission: Data centers will become "good citizens" of the grid, gently easing their power use in the rare hours the system is stressed, protecting your electricity bill and keeping the lights on for everyone else. But read what Nvidia said underneath the press release: Power has become "the defining constraint" on the expansion of AI infrastructure in the United States. The most valuable company in history, the one selling every chip the boom runs on, just told you the ceiling on its own business is the number of watts the country can physically deliver. So what is this alliance actually for? It's a lobbying group. And its real ask is this: Right now a new data center can wait a decade or more to connect to the grid, because utilities have to build for the worst hour of the year. The alliance wants governors and regulators to let any data center that promises to throttle its power during peak hours skip that line and plug in faster and bigger. And the grid is already buckling, and you're already paying for it. On PJM, the largest grid in America, serving 67 million people across 13 states, the wholesale cost of power jumped 76% in a single year. The independent market monitor pinned it directly on data centers and said the damage "is not reversible." Capacity prices in that market went from about $29 per megawatt-day to $329, roughly 11 times higher, and data centers drove most of it. This summer PJM had to beg households to cut their usage during a heat wave to avoid rolling blackouts. That's the backdrop the word "affordability" is doing so much work against. Now look at what the alliance actually brings to fix it: Google runs a demand-response program of roughly 1 gigawatt. Emerald and Nvidia have run 6 demonstrations. The first genuinely flexible AI data center they can point to is a single 100 megawatt site in Virginia that isn't even switched on yet. That's the pilot they're holding up against a buildout measured in tens of gigawatts, ordered years in advance, and already raising your rates today. So strip the civics off it and here's what actually happened: The best-funded companies on Earth just formed a coalition whose entire premise is that they need permission to sometimes use LESS electricity. They are announcing that the grid ran out of room before the models ran out of ambition, and they would like to jump the queue while you cover the upgrades. This has nothing to do with strength. Every earlier limit on AI was just a number on a slide, whether it was compute or funding or chip supply, and every one of them got solved with more money. This one can't be solved that way. You cannot venture-fund a power plant into existence in a quarter, and you cannot print a transmission line. Electricity is the first ceiling the AI industry can't buy its way through, and today it admitted that while dressing it up as a favor to your utility bill.

Ricardo

26,257 Aufrufe • vor 11 Tagen

Andrew Ng just revealed why the AI companies throwing the most compute at the problem are going to lose. The winner of the intelligence race won’t use the most compute. They’ll waste the least. Ng: “Most of your high-dimensional data lies on a lower-dimensional subspace. It’s just a fact of life.” Here’s what that means in practice. You have a 10,000-dimensional dataset. Every dimension dragged through every calculation. Every training cycle hauling dead weight the model will never use. Ng: “You’re carrying around these 10,000-dimensional examples throughout your whole training process.” That bloat isn’t just inefficient. It’s a tax on every computation you run. Memory bandwidth. Network bandwidth. Computational speed. All of it eaten by dimensions that contribute nothing to intelligence. They contribute noise. The insight that separates the architects from the arms race: that 10,000-dimensional dataset is almost entirely captured by a much smaller subspace. The signal lives in a fraction of the space you’re paying to process. Compress it. 10,000 dimensions down to 1,000. Ng: “You can run your learning algorithm on a much lower-dimensional set of data and it may be much more efficient.” Same hardware. Same budget. A fraction of the friction. Brute force is the strategy of whoever has the deepest pockets. Compression is the strategy of whoever actually understands the problem. The companies that master this don’t just build faster models. They build models that find more truth in less data than anything scaling blindly ever will. Intelligence was never about processing everything. It’s about knowing what to cut.

Dustin

215,643 Aufrufe • vor 7 Monaten

Demis Hassabis confirmed every frontier AI lab is working on recursive self-improvement and in the same sentence said the safety risk of removing humans from the loop entirely keeps him up at night. That combination should stop you. The CEO of Google DeepMind just confirmed that the thing most people treat as a theoretical future risk is already the active focus of every serious lab on earth right now. He explained why it works in coding and math. The feedback loop is fast. You can verify whether an answer is correct almost instantly. You can generate synthetic training data from it. The loop closes quickly and cleanly. Then he said where it breaks down. In biology, chemistry and physics. Any domain where verifying a hypothesis requires a physical experiment in the real world. The loop does not close in seconds. It closes in weeks or months. Geoffrey Hinton said in his Nobel lecture that recursive self-improvement is the development he fears most and that once started it may not be possible to stop. Hassabis is not pushing back on that. He is describing the guardrails labs are building around a process they are already running. Every lab has to think carefully about the safety of a process where no human is in the loop. He said that as a constraint they are navigating right now. The question they are sitting with is how much of it to let run without a human watching. (Watch the full interview on YouTube at Two Minute Papers channel)

Ihtesham Ali

68,231 Aufrufe • vor 3 Monaten

Greg Brockman, President of OpenAI, said there is not enough compute in the world to satisfy AI demand, and OpenAI itself cannot launch products it has already built because it cannot find the infrastructure to run them (Save this). OpenAI is spending $50 billion on compute in 2026 alone and it still is not enough. That is the setup but here is the trade. Nebius is one of the most asymmetric infrastructure plays in public markets right now, and most people have never heard of it. Q1 2026 revenue came in at $399 million, up 684% year over year, with AI cloud revenue specifically growing 841% in a single quarter. The company entered 2026 with an exit ARR of $1.25 billion and is targeting $7 to $9 billion by year end, a number that would make it one of the fastest revenue ramps in the history of public infrastructure companies. The contracted backlog sits at $50 billion anchored by a $17.4 billion agreement with Microsoft through 2031 and a $27 billion five-year deal with Meta. They are decade-scale infrastructure commitments from the two largest enterprise AI spenders on earth, signed before the demand curve has even reached its steepest point. Nvidia took a direct equity stake in Nebius, one of only two neoclouds it has invested in alongside CoreWeave. That relationship is not just financial but rather means Nebius gets preferential access to GPU allocation at a moment when every lab and every hyperscaler is competing for the same constrained supply. Contracted power capacity now exceeds 3.5 gigawatts, with expansion plans targeting 5 to 6 GW by mid-2029. And power is the other binding constraint in AI infrastructure, you cannot build a data center without it and Nebius has already secured the capacity that competitors are still fighting to acquire. At full ramp, analysts project revenue in the $15 to $25 billion range by 2029, against a current market cap the contracted backlog alone already dwarfs. Come join Milk Road Pro and get our full Nebius deep-dive, the exact price levels we are watching, how we are sizing the position against the backlog and power capacity timeline, and our full AI thesis. link below!

Milk Road AI

14,578 Aufrufe • vor 3 Monaten

If intelligence is the log of compute… it starts with a lot of compute! And that’s why we’re scaling our GPU fleet faster than anyone else. Just last year, we added over 2 gigawatts of new capacity – roughly the output of 2 nuclear power plants. And today we’re going further, announcing the world's most powerful AI datacenter, located in southeastern Wisconsin. Fairwater is a seamless cluster of hundreds of thousands of NVIDIA GB200s, connected by enough fiber to circle the Earth 4.5 times. It will deliver 10x the performance of the world’s fastest supercomputer today, enabling AI training and inference workloads at a level never before seen. For AI training workloads, you need compute at exponential scale. That’s why we designed the datacenter, GPU fleet, and network together as one integrated system. This ensures a single job can run from day 1 at exponential scale across thousands of GPUs. Fairwater uses a liquid-cooled closed-loop system for cooling GPUs that requires zero water for operations after construction. And we’re matching all of the energy that is consumed with renewable sources. And of course, it is just one of several similar sites we’re lighting up across our 70+ regions. We have multiple identical Fairwater datacenters under construction in other locations across the US, in addition to our AI infrastructure already deployed in over 100 datacenters around the world, powering model training, test-time compute, RL tuning, and real-time inference at global scale. Too often during times like this, people go with the current and only later wonder, how did we get here? With Fairwater, we're charting a new path: doing the hard engineering work, bringing compute, network, and storage into one highly scaled cluster, and designing closed-loop energy systems to meet real-world computing needs. And partnering with local communities to ensure it's thoughtfully done in a way that is sustainable, creates new jobs, and expands opportunity. We are thrilled to see this take hold in Wisconsin, and we are just getting started.

Satya Nadella

2,026,666 Aufrufe • vor 1 Jahr