Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

"Nvidia is positioned perfectly to thrive on the coding agent wave" and the explosion in inference demand, says tae kim. "I met with Ian Buck and dozens of engineers at Meta, Google, and Nvidia. All of them are seeing crazy inference demand and AI compute shortages." "People are building...

32,912 görüntüleme • 4 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

Micron is going to $4,000 and once you understand what inference actually is, the number stops sounding crazy (Save this). Dylan Patel just said that by 2030, OpenAI and Anthropic alone will need over 100 gigawatts of compute combined and by 2040, we may not even be measuring AI infrastructure in gigawatts anymore. We may be talking about terawatts. Every single one of those gigawatts needs memory to function. Without it, the compute is worthless. Most people heard that and thought about Nvidia but they should be thinking about Micron. Every AI model generating a response has two phases. The first is prefill, processing your prompt which is compute-heavy and the second is decode generating each word one token at a time and that phase is almost entirely memory-bound, not compute-bound. During decode, the GPU's processing units sit idle more than 95% of the time, waiting for data to arrive from memory. Google confirmed it in a research paper that decode-phase bottlenecks are dominated by memory bandwidth and capacity not raw compute. The GPU is not the bottleneck but the memory feeding the GPU is. This matters because inference is now where all the money lives. Training a model happens once, Inference happens billions of times a day every ChatGPT response, every Claude output, every agentic workflow running in the background and every one of those token streams is a billing event tied directly to memory performance. Adding more GPUs does not fix this because GPUs are already underutilized in inference because they are sitting idle waiting on memory. Adding more memory bandwidth and capacity is what directly reduces token cost, reduces latency, and allows the same cluster to serve dramatically more users simultaneously. Longer context windows compound the problem further, a model running a 1 million token context window requires dramatically more memory per session than a 10,000 token window, and every new model generation pushes context longer. The market treats memory as a downstream beneficiary of Nvidia orders. The correct framework is the opposite, Micron is the upstream constraint on how much value every Nvidia GPU can actually generate at inference scale. Micron guided Q4 to $50 billion in revenue, has HBM4 ramping at twice the pace of the prior generation, and CEO Sanjay Mehrotra has said supply will not catch demand before the end of 2027. At 8x forward earnings on $112 projected FY2027 EPS, Micron is the most undervalued infrastructure company in the entire AI stack. Inference is memory. Memory is Micron and the inference ramp has barely started. Milk Road Pro members are already up massively on this position and we're just getting started. If you want the full breakdown of what we're buying and why, come join us for just a dollar using the link below!

Milk Road AI

128,522 görüntüleme • 28 gün önce

Jensen Huang just admitted Nvidia is paying for BOTH ends of its own $500 billion deal. Every outlet ran the same headline: Nvidia is putting half a trillion dollars into Korea. The largest AI infrastructure commitment ever announced with a single partner. But when asked what was actually inside that number and whether this is Nvidia spending money in the Korean economy, or SK fronting the capital themselves, Jensen said this: "We're gonna be purchasing memories from them for many years to come. In order to build a trillion dollars worth of Vera Rubin systems, you're gonna have to buy a lot of system memories to go with it. And so we have large purchase agreements and large purchase intentions with SK Hynix. Meanwhile, SK Telecom is gonna become an AI cloud. And in that agreement, we will be selling AI supercomputers to them. So between us, we're gonna do $500 billion worth of business." He literally described two completely different transactions and added them together. Nvidia pays SK Hynix for memory chips. SK Telecom pays Nvidia for supercomputers. Both directions get stacked into one figure and handed to the market as demand. A large share of that half trillion dollars is Nvidia's own money going OUT the door. This is the CEO of the most important chip company on Earth. His silicon runs every serious AI system on the planet. If anyone alive could announce a clean half trillion in customer demand, it's him. Instead he announced a number that counts his own supplier payments as "business." But now this is where it gets genuinely crazy... Bloomberg asked how badly Nvidia needs Korean supply to grow. Jensen said Nvidia does not have enough bits. He said the company is constrained in HBM memory, constrained in LPDDR memory, and constrained in just about every part of the supply chain. Then he named the bottleneck nobody expected: "We're even constrained now with land and power and construction workers to set up the data centers." The company announcing the biggest AI deal in history cannot hire enough people to pour concrete. Then he capped the entire industry. He said the industry has the ability to double each year, and will have a hard time growing much faster than that. Now hold that against the target he set at the top of the same interview: He said the semiconductor industry probably has to become 10x larger than it is today over the next decade. Doubling annually clears that on a spreadsheet. But in reality it only happens if land, power and construction crews cooperate every single year for ten years straight. And the demand justifying all of it is a number most people have not heard yet. Jensen is planning for 100 BILLION AI agents and billions of robots using computers. That is the actual bet. 10x the industry, financed by deals where the vendor is also the customer, and gated by how fast you can find electricians. Korea's market did not celebrate any of this by the way. The KOSPI has been falling hard and SK Hynix and Samsung both slumped while the half trillion dollar headline was running. Jensen listed both directions of money and totaled them on camera because he does not think there is anything wrong with it. Maybe there isn't. Chips have to be bought before systems can be sold. But the market is pricing these announcements as demand, and at least half of this one is spending.

Ricardo

33,559 görüntüleme • 3 gün önce

Jensen Huang just made a statement that every investor in AI infrastructure needs to hear (Save this). He said that the AI buildout is accelerating, the second half of this year is going to be much larger than the first half, and next year is going to be very, very large. Micron is the best positioned to win from this because every Nvidia GPU requires High Bandwidth Memory stacked directly on the chip to feed it data fast enough to keep up. There is no AI compute without memory, and right now there is simply not enough memory to go around. Micron's entire HBM supply for 2026 is already completely sold out under multi-year agreements before the year even started. Micron's own management has acknowledged they can only satisfy 50 to 65 percent of demand from some of their most important customers. That is not a problem that gets fixed quickly, because new fabs take years to build. Micron's Idaho expansion does not come online until mid-2026, a second Idaho facility is not expected until 2028, and a new New York fab is looking at 2030. The demand Jensen just described is arriving right now, and the supply to meet it is years away. The financial results already reflect this dynamic. Micron's Q2 fiscal 2026 revenue came in at $23.86 billion, nearly triple what it was a year earlier beating consensus by roughly $3.8 billion. The HBM market alone is expected to grow from $35 billion today to $100 billion by 2028, and Micron has been consistently ahead of that forecast. Jensen just told the world the second half of this year and all of next year are going to be larger than anything that came before. Micron is the company that supplies the memory those GPUs need to run, and it cannot build supply fast enough to keep up with demand. Come join Milk Road Pro for our full deep dive on Micron, the HBM supply thesis and our AI trade thesis! Link below!

Milk Road AI

77,554 görüntüleme • 1 ay önce

Nebius will be the first neocloud to hit $1 trillion dollar company and here is exactly why (Save this). As dylan patel says Jensen Huang absolutely hates a world where the hyperscalers have all the power. A world where Microsoft, Amazon, and Google are the only ones building compute is a world where Nvidia is slowly being squeezed by a handful of customers all simultaneously developing custom chips to replace Nvidia GPUs entirely. Google's TPU, Amazon's Trainium and Microsoft's Maia all exist for one reason, to cut Nvidia out of the stack and Jensen knows it so he is playing a long game most investors haven't registered yet. By funding NeoClouds and NeoLabs at scale, Jensen is deliberately engineering a multipolar compute world where no single hyperscaler can dictate terms and where Nvidia hardware remains the default infrastructure layer regardless of which model or platform ultimately wins. Nvidia has deployed roughly $40 billion in AI ecosystem investments across OpenAI, Anthropic, CoreWeave, Nebius, xAI, and dozens of infrastructure companies, all running almost exclusively on Nvidia chips, cementing GPU dependency across the entire AI stack.sedaily Every neocloud that survives and scales becomes a permanent Nvidia GPU customer structurally opposed to the hyperscalers building custom silicon expanding Nvidia's market while simultaneously weakening its biggest competitive threat. Dylan Patel described the neocloud ecosystem as throwing bait into the water and letting the best fish survive, warning that many heavily-backed teams will fail, but the ones that emerge will pull hundreds of millions in ARR right out of the gate. Nebius is that fish because it's the only neocloud operating at hyperscaler scale while remaining fully purpose-engineered for AI workloads from silicon to software. The numbers confirm Nebius has already cleared the survival bar that will eliminate most of the 200+ neoclouds competing right now. Revenue hit $399 million in Q1 2026, up 684% year-over-year, backed by $46 billion in contracted backlog, 3.5 GW of contracted power across seven site and a target of $7–$9 billion in annualized revenue by year-end. When Google approached neoclouds about deploying TPUs, Nebius said no, its Chief Revenue Officer noting that demand is 99% for Nvidia GPUs and that TPU interest comes almost entirely from former Google employees rather than the actual market. That alignment with Nvidia's ecosystem, at this scale, with this backlog, and this level of strategic backing is why Nebius sits in a category of one among the neocloud field. Patel framed the broader play correctly, every neocloud that survives makes Google's TPU and Amazon's Trainium structurally weaker simply by existing and five years from now, the winners will have reshaped the entire compute landscape in Nvidia's favor. Nebius is already hundreds of millions in ARR ahead of the competition while most of the field is still treading water. Milk Road subscribers are already up massively on the Nebius trade, and we are tracking the neocloud buildout as Nvidia works to reshape the entire compute market. Come join Milk Road Pro for our full Nebius breakdown, the valuation framework, the revenue targets we are watching, and the AI infrastructure names we like next for just $1. Link below!

Milk Road AI

92,650 görüntüleme • 28 gün önce

Jensen Huang just doubled NVIDIA's demand forecast to $1 Trillion through 2027 🤯 Then spent two hours explaining why that number is conservative… Here's everything today from GTC: - NemoClaw: NVIDIA's open-source enterprise AI agent stack built around OpenClaw. Jensen called OpenClaw "the operating system for personal AI" and said every company needs a strategy for it. - Space-1: NVIDIA is putting Vera Rubin data centers in orbit. Not a concept. An actual system being designed for space deployment right now. - DLSS 5: 3D-guided neural rendering that blends raw graphics with generative AI. Jensen called it the future of real-time rendering. - AWS: Deploying 1 million+ NVIDIA GPUs starting this year. Azure was the first hyperscaler to power up Vera Rubin. - Vera Rubin: NVIDIA's next-gen AI supercomputer. 10x more performance per watt than Blackwell, 700 million tokens per second, shipping later this year. - Groq 3 LPU: First chip from NVIDIA's $20B Groq acquisition. A purpose-built inference accelerator that ships Q3. NVIDIA now owns training AND inference. -Feynman: The architecture after Rubin, coming 2028. New GPU, new LPU, new CPU. NVIDIA is on a 12-month chip cadence and the treadmill never stops. - Autonomous driving: BYD, Hyundai, Nissan, and Geely building Level 4 vehicles on NVIDIA. Uber deploying NVIDIA-powered robotaxis across 28 cities by 2028. The man doubled his demand forecast to a trillion dollars, announced data centers in space, and closed the show with a robot singing country music. This is NVIDIA's world. Everyone else is just renting compute in it.

Josh Kale

45,875 görüntüleme • 4 ay önce