Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Modular Expansion vs Vertical Integration Featuring: Logarithmic Rex rushi Nick White Kyle Samani Josh Tobkin (SUPRA)

37,244 görüntüleme • 1 yıl önce •via X (Twitter)

5 Yorum

Hoyt Dwyer profil fotoğrafı
Hoyt Dwyer1 yıl önce

@LogarithmicRex @rushimanche @nickwh8te @KyleSamani @JoshuaTobkin @JoshuaTobkin nails it. Vertical integration is the way.

Jon Jones | Supra profil fotoğrafı
Jon Jones | Supra1 yıl önce

@LogarithmicRex @rushimanche @nickwh8te @KyleSamani @JoshuaTobkin Integrated

Mystic Ledger profil fotoğrafı
Mystic Ledger1 yıl önce

@LogarithmicRex @rushimanche @nickwh8te @KyleSamani @JoshuaTobkin Permissionless, great debate topic. Modular expansion seems like a safer bet, but vertical integration can lead to crazy innovation - and I'm all for crazy

Ngọc Hải ꧁IP꧂ profil fotoğrafı
Ngọc Hải ꧁IP꧂1 yıl önce

@LogarithmicRex @rushimanche @nickwh8te @KyleSamani @JoshuaTobkin @SUPRA_Labs sẽ thay đổi cuộc chơi

Moses. D $SUPRA profil fotoğrafı
Moses. D $SUPRA1 yıl önce

@LogarithmicRex @rushimanche @nickwh8te @KyleSamani @JoshuaTobkin #supra the game changer🔥🔥🔥

Benzer Videolar

⚾️9/6 MLB MOST Notables⚾️ Show support for what we do👉 James Wood vs Justin Wrobleski: 2-2, 2 HRs💣💣, 108.9 EV Austin Riley vs Aaron Nola: 23-66, 5 Doubles, 6 HRs💣💣💣💣💣💣 Ronald Acuña Jr. vs Aaron Nola: 21-60, 6 Doubles, 4 HRs💣💣💣💣 Kyle Schwarber vs Tyler Mahle: 6-20, 2 HRs💣💣 Jose Altuve vs Eduardo Rodriguez: 10-30, 2 Doubles, 2 HRs💣💣 Yordan Alvarez vs Eduardo Rodriguez: 9-18, 2 Doubles Jeremy Peña vs Eduardo Rodriguez: 5-13, 1 Double Tyler Stephenson vs Kyle Harrison: 2-2, 2 HRs💣💣, 103 EV Brice Turang vs Brady Singer: 4-11, 3 Doubles Bo Naylor vs Brady Singer: 5-12, 2 Doubles, 1 HR💣 Matt Olson vs Aaron Nola: 10-38, 3 Doubles, 4 HRs💣💣💣💣 Mike Yastrzemski vs Aaron Nola: 5-15, 2 Doubles, 1 HR💣 Trea Turner vs Tyler Mahle: 8-19 Connor Wong vs Kyle Bradish: 2-5, 1 HR💣 Gunnar Henderson vs Payton Tolle: 3-6, 3 Doubles Christian Encarnacion-Strand vs Payton Tolle: 2-3, 1 HR💣 Javy Báez vs Gavin Williams: 5-8, 1 Double, 1 HR💣 Tyrone Taylor vs Tyler Phillips: 2-3, 1 HR💣 Nathan Lukes vs Randy Dobnak: 1-3, 1 HR💣 Carter Jensen vs Spencer Arrighetti: 2-4, 1 HR💣, 106.7 EV Ryan Vilade vs Mackenzie Gore: 1-3, 1 HR💣 Yandy Díaz vs Mackenzie Gore: 1-3, 1 HR💣 Nick Fortes vs Mackenzie Gore: 3-9, 2 Doubles Jake McCarthy vs Kyle Leahy: 3-3, 1 Double Darell Hernaiz vs Bryan Woo: 2-7, 1 HR💣 Jazz Chisholm Jr. vs Michael King: 2-5, 1 HR💣, 100 EV Jake Cronenworth vs Gerrit Cole: 1-3, 1 HR💣 Manny Machado vs Gerrit Cole: 3-3, 1 Double Royce Lewis vs José Urquidy (*Bulk): 2-5, 1 HR💣 Luke Keaschall vs José Urquidy (*Bulk): 2-2, 1 HR💣, 100.8 EV Bailey Ober vs Current White Sox: .190 BAA, 21.8% K-Rate🔥 🍓Fresh Matchups (never faced): Walbert Ureña vs PIT, Christian Scott vs SFG, Cesar Perdomo vs NYM, Andrew Alvarez vs LAD Who do YOU🫵 think adds to their notable history today/tonight?

Tablesetters: A Baseball Podcast

47,256 görüntüleme • 13 gün önce

⚾️6/29 MLB MOST Notables⚾️ Casey Schmitt vs Eduardo Rodriguez: 5-8, 2 HRs💣💣 Matt Chapman vs Eduardo Rodriguez: 8-24, 4 Doubles, 1 HR💣 Ketel Marte vs Tyler Mahle: 8-15, 3 Doubles, 2 HRs💣💣 Willson Contreras vs Miles Mikolas: 7-19, 2 Doubles, 1 HR💣 Caleb Durbin vs Miles Mikolas: 4-7, 2 Doubles, 1 HR💣 Marcell Ozuna vs Aaron Nola: 17-70, 1 Double, 5 HRs💣💣💣💣💣 Miguel Vargas vs Shane Baz: 1-2, 1 HR💣 Taylor Ward vs Sean Burke: 4-9, 2 Doubles Sean Burke vs Current Orioles: 6-39, 11 K🔥 Brandon Lowe vs Aaron Nola: 2-7, 1 HR💣, 106.6 EV Endy Rodriguez vs Aaron Nola: 2-6, 1 Double, 1 Triple💨 Kyle Schwarber vs Braxton Ashcraft: 2-3, 1 HR💣 Bryce Harper vs Braxton Ashcraft: 3-3 Ryan Weathers vs Current Tigers: 7-37, 11 K🔥 Jazz Chisholm Jr vs Casey Mize: 4-10, 1 HR💣 George Springer vs Sean Manaea: 13-41, 2 Doubles, 1 HR💣 Ranger Suarez vs Current Nationals: .156 BAA, 30.4% K-Rate🔥 Jacob Young vs Ranger Suarez: 3-5, 1 Double Nick Lodolo vs Current Brewers: .182 BAA, 22.5% K-Rate🔥 Shota Imanaga vs Current Padres: .184 BAA, 26.4% K-Rate🔥 Josh Bell vs Peter Lambert: 3-5, 1 HR💣 Isaac Paredes vs Zebby Matthews: 1-3, 1 HR💣 Max Muncy vs Eric Lauer: 1-2, 1 HR💣, 105.6 EV George Kirby vs Current Angels: .194 BAA, 32.5% K-Rate🔥 🍓Fresh Matchups (never faced): Trey Yesavage vs NYM, Gage Jump vs LAD Who do YOU🫵 think adds to their notable history today/tonight?

Tablesetters: A Baseball Podcast

81,970 görüntüleme • 2 ay önce

⚾️6/5 MLB MOST Notables⚾️ Ronald Acuña Jr vs Mitch Keller: 8-19, 3 HRs💣💣💣 Matt Olson vs Mitch Keller: 6-12, 1 Double, 1 HR💣 Alec Burleson vs Brady Singer: 3-7, 2 Doubles, 1 HR💣 Masyn Winn vs Brady Singer: 2-5, 1 Double, 1 HR💣 Michael Massey vs Zebby Matthews: 5-8, 1 Double, 1 HR💣 Jung Hoo Lee vs Edward Cabrera: 3-5 Luis Arraez vs Edward Cabrera: 2-6, 1 Double Robbie Ray vs Current Cubs: 8-44, 11 K🔥 Randy Arozarena vs Framber Valdez: 5-18, 1 Double, 2 Triples, 1 HR💣 Gleyber Torres vs Bryan Woo: 4-8, 1 Double Jesús Luzardo vs Current White Sox: 4-30, 6 K🔥 Brandon Marsh vs Anthony Kay: 1-2, 1 HR💣 Willson Contreras vs Ryan Weathers: 3-9, 2 Doubles, 1 HR💣 Cody Bellinger vs Sonny Gray: 3-10, 1 Double Jazz Chisholm Jr vs Sonny Gray: 3-7, 1 Triple💨 Trent Grisham vs Sonny Gray: 4-7, 1 Double Ben Rice vs Sonny Gray: 1-2, 1 HR💣 Jesús Sánchez vs Brandon Young: 5-12, 2 Doubles Ernie Clement vs Brandon Young: 4-6 Kyle Stowers vs Drew Rasmussen: 2-5, 1 Double Xavier Edwards vs Drew Rasmussen: 3-6 Jakob Marsee vs Drew Rasmussen: 2-3, 1 Double Marcell Ozuna vs Martín Pérez: 4-7, 3 Doubles Austin Riley vs Mitch Keller: 4-10, 1 Triple💨 Mauricio Dubón vs Mitch Keller: 4-10, 1 Double Nathaniel Lowe vs Kyle Leahy: 3-6, 1 HR💣 Josh Bell vs Michael Wacha: 6-26, 2 HRs💣💣 Christian Yelich vs Ryan Feltner: 5-12, 3 Doubles Luis García Jr vs Merrill Kelly: 3-6, 1 HR💣 James Wood vs Merrill Kelly: 1-3, 1 HR💣 MJ Melendez vs Michael King: 2-6, 1 HR💣 Bo Bichette vs Michael King: 6-21, 1 Double Mike Trout vs Roki Sasaki: 1-3, 1 Double 🍓Fresh Matchups (never faced): Kumar Rocket vs CLE, Parker Messick vs TEX, Foster Griffin vs ARI, Christian Scott vs SD Who do you think adds to their notable history today/tonight?

Tablesetters: A Baseball Podcast

173,983 görüntüleme • 3 ay önce

⚾️9/4 MLB MOST Notables⚾️ If you appreciate what we do and want to show a little support, anything is appreciated!!!👇 Jazz Chisholm Jr. vs Walker Buehler: 6-10, 1 Double, 2 HRs💣💣 Jackson Merrill vs Max Fried: 4-7, 1 Triple, 1 HR💣 Jake Bauers vs Rhett Lowder: 5-8, 2 HRs💣💣 Lane Thomas vs Cristopher Sánchez: 8-14, 2 Doubles, 1 HR💣 Spencer Torkelson vs Logan Allen: 5-8, 2 Doubles, 2 HRs💣💣 Zach McKinstry vs Logan Allen: 1-3, 1 HR💣 Gleyber Torres vs Foster Griffin: 1-3, 1 HR💣 Travis Bazzana vs Keider Montero: 1-3, 1 HR💣 Keider Montero vs Current Guardians: 10-69, 9 K🔥 Sal Frelick vs Rhett Lowder: 4-8, 1 Double William Contreras vs Rhett Lowder: 4-8, 1 HR💣 Eugenio Suárez vs Shane Drohan: 2-4, 2 Doubles Matt Olson vs Cristopher Sánchez: 6-20, 2 Doubles, 2 HRs💣💣 Ronald Acuña Jr. vs Cristopher Sánchez: 4-13, 1 HR💣 Ozzie Albies vs Cristopher Sánchez: 8-15, 2 Doubles Austin Riley vs Cristopher Sánchez: 6-12, 1 HR💣 Trea Turner vs Chris Sale: 6-19, 2 Doubles Kyle Schwarber vs Chris Sale: 5-19, 1 Triple, 2 HRs💣💣 Shane Baz vs Current Red Sox: .198 BAA, 26.3% K-Rate🔥 Coby Mayo vs Ranger Suarez: 2-3, 1 HR💣 Erick Fedde vs Current Twins: .151 BAA, 25% K-Rate🔥 Drew Romo vs Zebby Matthews: 2-5, 1 HR💣 Zebby Matthews vs Current White Sox: .177 BAA, 19.3% K-Rate🔥 Joc Pederson vs Nick Martinez: 2-5, 2 HRs💣💣, 105.7 EV Elias Díaz vs Nick Martinez: 4-8, 1 Double, 1 HR💣 Nolan Arenado vs Cristian Javier: 2-5, 2 HRs💣💣 Vladimir Guerrero Jr. vs Daniel Lynch IV: 4-8, 1 HR💣 Ernie Clement vs Daniel Lynch IV: 3-6, 1 HR💣 Bobby Witt Jr. vs Jameson Taillon: 2-3, 1 HR💣 Iván Herrera vs Ryan Feltner: 2-8, 1 HR💣 Fernando Tatis Jr. vs Max Fried: 5-10, 1 Double Austin Hays vs Max Fried: 5-9, 1 Double Zack Gelof vs Logan Gilbert: 3-8, 1 Double, 1 HR💣 Carlos Cortes vs Logan Gilbert: 3-5, 100.0 EV 🍓Fresh Matchups (never faced): Andrew Sears vs CLE, Matt Wilkinson vs NYM, Jackson Kent vs LAD, Kade Morris vs SEA Who do YOU🫵 think adds to their notable history today/tonight?

Tablesetters: A Baseball Podcast

47,149 görüntüleme • 15 gün önce

⚾️6/24 MLB MOST Notables⚾️ Seiya Suzuki vs Nolan McLean: 2-3, 2 HRs💣💣 Brandon Nimmo vs Eury Pérez: 3-5, 1 HR💣 Brett Baty vs Shota Imanaga: 2-5, 1 HR💣 Francisco Alvarez vs Shota Imanaga: 1-2, 1 HR💣 Dansby Swanson vs Nolan McLean: 1-1, 1 HR💣 Michael Busch vs Sean Manaea: 2-5, 1 HR💣 Paul Goldschmidt vs Tarik Skubal: 5-10, 2 HRs💣💣 Steven Kwan vs Erick Fedde: 5-9, 1 HR💣 Rhys Hoskins vs Erick Fedde: 10-36, 1 Double, 3 HRs💣💣💣 Daniel Schneemann vs Erick Fedde: 3-6, 2 Doubles, 1 Triple💨 Tanner Bibee vs Current White Sox: 6-38, 10 K🔥 Andruw Monasterio vs Kyle Freeland: 2-5, 1 HR💣 Tyler O’Neill vs José Soriano: 1-1, 1 HR💣 Nick Gonzales vs Bryan Woo: 3-3, 1 Double Tarik Skubal vs Current Yankees: .167 BAA, 27% K-Rate🔥 Bryce Harper vs Miles Mikolas: 9-25, 2 Doubles, 2 HRs💣💣 Alec Bohm vs Miles Mikolas: 7-16, 2 Doubles, 1 HR💣 Edmundo Sosa vs Miles Mikolas: 1-2, 1 HR💣, 110 EV Luis García Jr vs Aaron Nola: 13-41, 3 Doubles, 2 HRs💣💣 Dylan Crews vs Aaron Nola: 5-10, 1 Double, 1 Triple💨 Max Muncy vs Joe Ryan: 1-3, 1 HR💣 Byron Buxton vs Shohei Ohtani: 1-2, 1 HR💣 Mauricio Dubón vs JP Sears: 5-17, 2 Doubles, 1 HR💣 Manny Machado vs Martín Pérez: 7-20, 1 Double, 1 HR💣 Tyler Soderstrom vs Tyler Mahle: 3-6, 1 Double 🍓Fresh Matchups (never faced): Trey Gibson vs LAA, Trey Yesavage vs HOU, Shane Drohan vs CIN, Mitch Bratt vs STL, Gage Jump vs SFG Who do YOU🫵 think adds to their notable history today/tonight?

Tablesetters: A Baseball Podcast

62,175 görüntüleme • 2 ay önce

BREAKING: How AI Reaches the Most Extreme Parts of the World — From U.S. Navy Ships & Saudi Deserts to the Alaskan Arctic Dan Wright (Dan Wright), CEO of Armada Recorded at the Reagan National Defense Forum Dan joins Sourcery to share why the global race for AI sovereignty will shape the future of power And how Armada partners with SpaceX, OpenAI, Saudi Aramco, U.S. Navy, MSFT & more, to reach every edge Key Points - Starlink, Starshield, SpaceX IPO, data centers in space - 70% of the world does not have access to AI, Armada changes that - Modular data centers vs hyperscaler data centers - U.S. vs China AI stack - American AI dominance Armada has $200M+ in total funding from top investors like Founders Fund, Dragon Global, Lux Capital, Shield Capital, 8090 Industries, Microsoft’s M12, Overmatch, Silent Ventures, Felicis, & Marlinspike With more recent strategic participation from investors: Pinegrove, Veriten, & Glade Brook Armada’s partners include Microsoft, Starlink, OpenAI, NVIDIA, Nokia, Red Hat, Carahsoft, AVEVA, Esri, Halliburton, Skydio, & more Ronald Reagan Presidential Foundation & Institute HIghlights (00:00) Dan Wright, CEO Armada (01:12) Armada: an edge hyperscaler deploying modular AI data centers (04:28) Partnering with SpaceX, Microsoft, Skydio, and OpenAI (05:47) Deploying AI in extreme environments: Ocean, Desert, Arctic & Space (08:11) SpaceX IPO rumors and Starlink’s rapid expansion (12:10) The AI dominance white paper and US vs China infrastructure (14:54) Edge compute, latency, and real world response at scale (17:43) Armada’s roadmap for 2026

Molly O’Shea

16,376 görüntüleme • 7 ay önce

$HIMS| Adjustment on Growth toward 2030🧵 Not Financial Advice! FY2025: Revenue: $2.35B or 58% YoY (weightloss $740), Core $1.61B FY2026: Revenue $3.2B(36%) where weightloss may be down by 10-15% or flat. FY2027: Revenue $4.16B(30%) FY2028: Revenue: $5.2B(25%) FY2029 Revenue: $6.5B(25%) FY2030 Revenue: $8.12B(25%) I expect management to ramp up buyback from FCF generation while company is trading at under 2x P/S andrewdudum. The discontinuation of Hims & Hers' compounded oral semaglutide pill in early February 2026(after 2 days), prompted by FDA regulatory actions and legal pressures from Novo Nordisk, introduces near-term challenges to the weight loss segment but does not derail the company's broader growth trajectory, as it pivots aggressively toward diversification and high-potential expansions The weight loss category bolstered by liraglutide injectables, generic semaglutide in Canada, and non-GLP-1 personalized kits retains strong momentum, contributing approximately 31% of total revenue in 2025 and projected to grow at 15-20% annually through 2030, down from prior 60%+ rates but still adding $150-250 million yearly through cross-selling and retention. Offsetting this moderation are ambitious new expansions: international markets, now accounting for an initial 5-10% of revenue but scaling to 20% by 2030 via Canada entry (projected 10% growth contribution in 2026 from generic semaglutide and Livewell acquisition) and Europe/UK via Zava (adding 8-12% incremental growth through telehealth in Germany, France, and Ireland); diagnostics and labs, launched in late 2025 with Quest Diagnostics partnership and YourBio Health's pain-free blood sampling tech, offering 50-120 biomarker tests across heart, metabolism, hormones, inflammation, and stress, expected to generate 12-18% of total revenue by 2027 and ramp to a standalone $1 billion segment by 2030. Preventive care and longevity initiatives, set for full 2026 rollout including peptide manufacturing (via acquired U.S. facility, contributing 10-15% to growth through vertical integration and supply control), coenzymes, GLP/GIP blends for performance and recovery, and a $325 million Grail investment enabling multi-cancer early detection blood tests (projected to add 8-12% revenue uplift starting in 2026 by enhancing subscription retention); and hormone health expansions like menopause/perimenopause and low testosterone treatments, already driving 10% of 2025 growth and poised for 20-25% annual expansion through data-driven personalization. Multi-cancer early detection (MCED) blood testing via the Galleri® test from GRAIL in the prior breakdown, even though it was bundled under longevity/preventive care. This is a significant new offering launched on February 4, 2026, providing subscribers (via the Labs platform) access to a simple annual blood test that screens for signals shared by over 50 types of cancer (including hard-to-detect ones like pancreatic, liver, ovarian, and lung) before symptoms appear. Hims & Hers is offering it at a discounted ~$700 (vs. retail $949), following their participation in GRAIL's $325 million private placement investment in late 2025, which strengthens the partnership and positions this as a core pillar of proactive/longevity care. This could help push Average growth to 30-35% vs 28.2%(my above revised projection). These levers, combined with a subscriber base exceeding 2.5 million (up 31% YoY) and AI-enhanced platform efficiency under new CTO leadership, support an upward revision to growth rates targeting 22-25% CAGR from 2026-2030 to meet the company's $6.5 billion revenue goal, far outpacing prior conservative estimates of mid-teens expansion. This high-growth scenario assumes execution on global scaling, regulatory navigation (FDA approvals for compounded alternatives), and margin recovery to 74-78% via vertical integration, positioning Hims & Hers as a comprehensive digital health ecosystem rather than a GLP-1-dependent player, with potential upside from emerging trends like peptide demand (up 144% in Google searches) and proactive wellness adoption. Not Financial Advice!

Mike

273,519 görüntüleme • 7 ay önce

> celestia's matcha upgrade coming next monday: the first DA layer ready for nasdaq scale onchain finance. what did Nick White say tonight on The Rollup? ⇾ Celestia identified clob exchanges as the strongest product fit for its architecture, since exchanges need ultra low latency and very high throughput. single sequencer rollups provide the fastest latency while still enabling censorship resistance and non custodialness. ⇾ fully onchain clobs require posting all orders and cancels on chain, demanding mb/sec da capacity. the upcoming matcha upgrade enables 5 mb/s, and governance can scale this to 20 mb/s, enough for nasdaq level traffic. ⇾ teams like RISE are ready to use this throughput, reinforcing celestia's push toward vertical integration. ⇾ also depends on matcha's higher throughput, adding urgency to the upgrade. matcha reduces issuance to 2.5% and ships components of the long term proof of governance model, which removes staking and pays ~0.25% issuance directly to validators. ⇾ celestia is shifting from a build whatever approach toward targeting apps with real pmf, mirroring broader industry moves toward more opinionated architectures. because mass adoption is still far away, path dependency matters: the team believes only about 10 apps truly matter, and celestia needs to win some of them. ⇾ high volume markets like onchain forex require massive DA and celestia is one of the few platforms capable of supporting this scale. success of bullet or other clobs would prove that wall street scale onchain finance is possible. ⇾ tradfi teams are increasingly exploring building on chain platforms, and celestia must balance institutional needs with fast moving native teams. ⇾ private blockspace is already live in production with Hibachi with clear institutional and dex interest. celestia focuses mainly on app acquisition while dedicating ~20% to internal tooling.

Mora 🦥

31,824 görüntüleme • 10 ay önce

Law firms are putting AI in the wrong place. Sullivan & Cromwell, Latham, Allen & Overy - every major firm is racing AI into legal research, drafting, and memos. That's exactly where hallucinations become malpractice. A single fabricated case citation has already sanctioned real lawyers (Mata v. Avianca, 2023 - the ChatGPT lawyers). A hallucinated statute in client advice is worse. Meanwhile the one place AI is genuinely safe - intake, qualification, and scheduling still runs on PDF questionnaires and paralegal phone tag at almost every firm in the country. So last night I built what lawyers should actually be building. A demo website for a fictional U.S. immigration firm - Sterling & Reed, lead partner Ann Sterling (all names are fictional). An AI intake concierge named Evelyn qualifies every prospect through 17 consultative questions, books the consultation on Ann's Calendly, and emails a two-page matter brief straight to her Gmail before she joins the call. No briefs. No citations. No advice. No hallucination surface. Any immigration lawyer on this app can replicate it. Here are the 12 exact prompts I used - copy-paste into Claude Code: ━━━━━━━━━━━━━━━━━━━━ 1/ BRAND FOUNDATION "Design a boutique U.S. immigration firm website. Fictional founding partner, two offices (NY + Miami). Palette: deep navy + bronze + warm paper. Fonts: Instrument Serif for display, Inter for body. Luxury + editorial - no generic templates, no blue-chip blue." 2/ HERO "Full-screen dark cinematic hero. Centered serif headline: 'Your immigration lawyer, already [prepared].' The last word cycles every 3.5s, character-by-character morph - rotating through prepared / briefed / engaged / on your side. Background: 6 US city night-skyline photos crossfading every 5s with Ken Burns drift. Horizon glow + starfield overlay." 3/ AI INTAKE CONCIERGE "Build Evelyn, an AI intake concierge. 17-turn immigration intake: greeting → visa pathway → citizenship + status + expiration → professional background → timeline → visa-specific qualifier (EB-5 capital, O-1 evidence, E-2 treaty, etc.) → source of funds / sponsor → prior visa history + derivatives → red flag on prior denials → red flag on arrests / overstays → biggest concern → referral source → name → email → WhatsApp → present 3 slots → booking confirmation. Voice: warm, consultative, never rushed. Frame red flags as 'no wrong answers - Ann prefers to know upfront.'" 4/ THINKING STATE "Before each Evelyn reply, show a thinking state. Spinning bronze ring + context-aware label per turn ('Identifying visa pathway...' / 'Cross-referencing denials database...' / 'Preparing brief to Ann...'). Then typing dots. Then the reply. Feels deliberate, not robotic." 5/ AGENT AVATAR "Evelyn's avatar: real photo of a professional woman in a circle. Bronze conic-gradient ring rotating around her, sonar pulse rings expanding outward, green live-status dot bottom-right. Three states synced to chat activity: idle (gentle breathing), thinking (faster pulse + bronze glow halo), speaking (bronze waveform bars below photo)." 6/ BOOKING - CALENDLY INTEGRATION "After intake completes, embed the firm's Calendly inline in the chat for slot selection. On confirmation, show an animated card: 30-particle bronze burst + 4 cascading checkmarks 300ms apart - Brief delivered to Ann's Gmail → Calendar dispatched via Calendly → WhatsApp queued → Prep checklist attached." 7/ HOW IT WORKS - SCROLL-PINNED "4-step section pinned with GSAP ScrollTrigger: 01 Intake, 02 Routing, 03 Consultation, 04 Engagement. Each step: custom animated SVG (chat dots pulsing / checkmarks drawing / calendar slot pulsing / signature stroke drawing itself). As you scroll, active step scales up + glows, inactive steps dim + blur. Bronze progress bar fills the bottom of the active step." 8/ LIVE STAT BAND "One horizontal line: '1,247 Matters filed | 38 Countries of origin | 97% Approval rate.' White italic Instrument Serif numbers, bronze vertical rules between. On scroll-in: scramble-resolve animation over 1.5s. First stat then becomes a live ticker - every 10-24s increments by 1 with champagne flash + floating '+1 EB-5' / '+1 O-1' / '+1 Family' badge (weighted random matter type)." 9/ BEFORE vs AFTER "Editorial band showing '21 days → 6 minutes.' Left: huge italic serif '21 days' with diagonal strike-through that draws in on scroll + five struck-through bullets (12-page PDF, five emails, paralegal screening, conflicts memo, partner hand-off). Arrow. Right: italic bronze '6 minutes' + five clean bullets. Below: live session clock + three real-time counters (briefs filed, conflict checks cleared, calendar holds reserved) ticking up while the visitor reads." 10/ EDITORIAL TESTIMONIAL "Pull quote block. 200px italic bronze opening mark (❝) fades in at 18% opacity. Two-line quote with 300ms staggered reveal. Bronze underline draws under emphasized phrases. Below: bronze divider + initials circle + name + verified green pill ('● Verified client · 2025'). Bronze corner brackets top-left and bottom-right." 11/ REPRESENTATIVE WORK "3 recent matters as a vertical bronze timeline. On scroll, the line draws top-to-bottom and marker dots pop in with staggered sonar rings. Per matter: visa tag (EB-5 / O-1 / E-2), matter number ('No. 1,247'), italicized key figures, green outcome pill (I-526E Approved / Premium Processing Approved / First-Interview Approved)." 12/ BLOG + CONTACT + FLOATING BUTTONS "3 blog cards (Instrument Serif italic titles, bronze gradient placeholders): EB-5, O-1, Family-based. Simple contact form: name + email + country of citizenship + visa type + note. Dark footer with both offices. Two floating FABs: WhatsApp bottom-right (green sonar pulse, pre-filled message) + music toggle bottom-left." ━━━━━━━━━━━━━━━━━━━━ Built entirely in Claude Code. No Cursor, no React boilerplate, no design team. The intake bot runs as a deterministic server flow - no AI inference during the conversation itself, which is why it can't hallucinate. Briefs pipe to Gmail. Consultations book through Calendly. Deployed on Vercel in 15 minutes. Every tool a lawyer needs for this is either free or already in the firm. Swap the fictional firm for your name, your practice areas, your matters - customize and you're live by the weekend. The AI sits in intake, not in your brief. No hallucination, no malpractice, no sanction risk. Just a qualified lead, a warmed prospect, and a partner who walks into the consultation already prepared.

Ann Srivastava

18,982 görüntüleme • 4 ay önce

$AMD| The FOMO to buy AMD Chips is NOW 🧵 Not Financial Advice! DYOR! Research Purpose Only! The Inference Queen is the biggest winner in Agentic AI where all other CPUs are struggling to compete with a 2yr old EPYC Turin and EPYC Venice is in mass production phase. AMD stresses deployability today on standard x86 platforms (no proprietary architectures required), full software compatibility, and open standards. This positions Venice + Helios as a practical, high-density alternative to competing solutions while underscoring that agentic AI shifts the balance toward CPU-rich racks alongside GPUs, and most importantly, lowering the cost of token to accelerate adoption and innovation. Context: The Wall Street Journal yesterday came out with an article that OpenAI is condiering drasstically lowering the token prices to win more customers from Anthropic. The narrative "they" are trying to exacerbate the current AI selloff won't last long. This is a fundamental misunderstanding of what is going on, or what I already discussed for months and years. Followers and Subscribers already knew this for years, that this day would come, where token cost will bcome the central discussion among enterprises as there is no such thing as unlimited budget or Tokenmaxxing when they use $NVDA chips or In-house Hyperscalers chips. I will link various threads if you are interested in understanding the full picture from supply chain to recent TSMC Rapid 2nm expansion up to 12 Fabs total by 2027/2028. Hyperscalers and AI natives effectively have no choice but to buy more AMD system for Agentic AI as leadership in economical, power-aware, high-volume internal + agentic use. However, due to supply constraints where Supply is far behind Demand, this makes multi-vendor reality along with in-house chips drive faster industry progress, lower overall costs, and better sustainability. NVIDIA’s Vera Rubin cannot compete with a 2 years old EPYC Turin, but AMD under Dr. Lisa Su has engineered the lowest cost-per-million-tokens, highly competitive energy-efficient solutions, and superior CPU orchestration for agentic AI at scale with Helios. Dr. Su has championed this shift since at least 2023, foreseeing the rise of agentic workflows that demand far more orchestration, parallel agents, and balanced compute well before the industry fully embraced it. Her long-term vision of AI moving from simple prompts to always on, multi-agent systems has driven AMD’s investments in high-core EPYC CPUs and integrated rack-scale solutions, perfectly positioning the company for today’s realities. The OpenAI-AMD 1GW Helios deployment (starting H2 2026) represents a pivotal vertical integration move that directly supercharges the inference economics. This isn't incremental; it's a structural shift toward ownership of massive, optimized rack-scale capacity, enabling the lowest token costs and triggering the enterprise adoption flywheel. We need to be honest, $AMD is the only company that made a big bet on Inference since the day Chatgpt became sensational where $NVDA and others were betting big on Training. At the end of the day, Token bill from Anthropic has to obey economics. Meaning the bills rise, companies have to get more out of it to justify the cost. It cannot be an unlimited inference budget, and it has to show up on efficiency, profitability and operating leverage. 1. Tokenomics After you understand this, you will understand why Citi cited Anthropic is likely to sign a deal with $AMD along with Hyperscalers, AI Labs, Sovereign AI like Softbank 5GW in France and many other countries. However, OpenAI and $META are now wanting faster deployment, and they are AMD shareholders now, they have prioritized allocation. Anthropic and Hyperscalers just cannot compete when Helios Rack lower token cost to$0.0003–$0.0005 per million tokens at GW scale. Cost to build 1GW data center 1GW Helios Rack full build is estimated $30-$35B 1GW Rubin Rack full build is estimated $45-$55B Inference (Cost per Million Tokens) ~$NVDA B200 / HGX: ~$0.02–$0.08 on optimized workloads (FP4/MXFP4, speculative decoding). Significant improvement over Hopper but still premium-priced. GB200 NVL72 rack-scale: $0.05–$0.25+ ~$AMD Helios Racks: $0.0003-$0.0005 per M tokens, dramatically lower than NVIDIA equivalents in owned infra. MI355X node-level: Up to 40% more tokens per dollar vs. competing solutions ( B200), driven by higher memory capacity (up to 288GB+ HBM), strong bandwidth, and lower acquisition costs. Training ~$NVDA Rubin Rack is estimated $0.7-$1.2/M Tokens ~$AMD Helios Rack is estimated $0.65-$1.0/M Tokens Now, OpenAI, META and Hyperscalers can lower Inference cost even further with $AMD EPYC Venice "dense rack" or Agentic AI Rack. AMD published a detailed technical blog emphasizing that the future of agentic AI autonomous, multi-step AI systems requiring heavy orchestration, databases, caching, APIs, and control planes demands massive CPU-dense rack-scale infrastructure, not just GPUs. The catalyst prominently positions their upcoming 6th Gen EPYC "Venice" processors as the key enabler for next-generation dense racks, delivering leadership throughput under real-world power, cooling, and density constraints. ~EPYC Venice (Zen 6 architecture, up to 256 cores / 512 threads per socket) is projected to deliver exceptional rack-level performance. In AMD’s modeled 100 kW rack comparisons, Venice-powered systems are expected to achieve ~3.30x the throughput of NVIDIA’s Vera (88-core Olympus) baseline across a broad mix of agentic-supporting workloads. ~This builds on current-generation 5th Gen EPYC "Turin" (up to 192 cores), which already delivers ~2.37x rack throughput vs. Vera and ~1.6x vs. Intel’s Xeon 6980P (128 cores). ~ Liquid-cooled Turin deployments already support >27,000 CPU cores per rack today. Venice is architected to push this beyond 36,000 cores in the same rack class, dramatically increasing concurrent agent capacity and overall infrastructure efficiency. 2. Ownership vs renting compute from Hyperscalers matter to OpenAI and only owning $AMD chips can meaningfully lower token cost for enterprises. ~Eliminates cloud overhead: No provider margins, utilization buffers, or egress fees. Direct control over power contracts, cooling, scheduling, and orchestration at dedicated facilities. ~Helios optimizations at GW scale: Rack-level density (1.4+ exaFLOPS FP8 per rack), high HBM4 bandwidth, EPYC orchestration for agentic workloads, and superior TCO/TDP. AMD's long-standing focus on tokens per dollar/watt shines here 20-40%+ efficiency edges in inference-heavy scenarios. ~At 1GW+ optimized deployment, inference hits $0.0003–$0.0005 per million tokens (community/analyst models tied to Helios metrics). This is dramatically lower than typical rented/cloud equivalents, especially for high-volume output tokens in agentic flows. High token bills today, enterprises running heavy agentic/coding/analysis workloads can face $50-100M+/month at current API rates (flagship models $5-30+/M output, scaled to massive volumes). Post-Helios compression, same volume will drop to $10-15M/month (or better) via lower underlying costs passed through as pricing flexibility, volume tiers, caching, or batch discounts. ROI thresholds collapse. More companies greenlight pilots → production → massive scaling. Agentic AI (autonomous workflows) multiplies token demand exponentially, but affordability removes the friction. OpenAI gains flexibility, Unlike more cloud-dependent rivals (Anthropic), they can lower effective pricing, offer aggressive enterprise bundles, or absorb volume without margin destruction directly tackling "high token bill" complaints while maintaining profitability as usage explodes. 3. Agentic AI Models shifted CPU:GPU Ratio to 1:1 toward 3-5:1 with Explosively Token-Hungry Workloads Agentic AI (autonomous, multi-step agents with planning, tool use, iteration, and self-correction) is fundamentally more compute and token intensive than conversational or single-turn generative AI. Agentic AI. autonomous, multi-step workflows with orchestration, tool use, parallel agents, data movement, and enterprise integration has dramatically increased the importance of strong host CPUs alongside GPUs. This shifts the CPU-to-GPU ratio higher and makes balanced systems critical toward 1:1 to 5:1 as enterprises testing more than 5-10 agents. AMD EPYC Venice excels ~Leadership core density (up to 256 Zen 6 cores per socket) for running many agents in parallel, orchestration layers, and high-throughput control-plane tasks. ~Superior performance-per-core and power efficiency ( up to 2.1x higher perf/core and 2.26x better SPECpower vs. NVIDIA Grace in benchmarks). ~Tight integration in Helios: One Venice CPU + multiple MI450 GPUs per node, enabling efficient data feeding to GPUs ("zero-copy"), parallel execution, and full rack utilization for complex agentic loops. Hyperscalers (Meta, Microsoft, Amazon, Google, Softbank) and AI natives (OpenAI, Anthropic...) are adopting high-core EPYC at scale specifically for these agentic demands, as CPUs now handle a larger share of non-model work (orchestration, policy enforcement, tool calls). This complements AMD’s lower-cost GPUs for overall TCO wins. ~Agents often generate 10–100x+ more tokens per task due to iterative reasoning chains, multiple tool calls, verification loops, and long-context orchestration. ~Goldman Sachs forecasts token consumption multiplying 24x by 2030 (to 120 quadrillion tokens/month) largely driven by agentic adoption in consumer and enterprise. ~Enterprise data shows agent-pattern workloads growing at 680% annualized rates, projected to surpass conversational AI in token volume by Q3 2026. ~Daily enterprise agent token consumption is already in the billions, with complex workflows (coding, workflows, analysis) amplifying this dramatically. 4. Competitive Edge: Winning Customers from Anthropic Anthropic’s Claude models (especially Opus/Sonnet) excel in complex reasoning and agentic coding, commanding premium positioning. However, their higher underlying costs (heavier reliance on third-party cloud with margins) limit pricing flexibility compared to OpenAI’s owned Helios capacity. Anthropic is on track to generate $10.9 billion in Q2 revenue. The company expects to achieve its first-ever quarterly adjusted operating profit of $559 million. However, sustaining full-year profitability remains challenging due to immense computing and model training costs The truth is, Anthropic has no choice but to buy as much $AMD chips as possible if they want to compete with OpenAI or get investors attention. This 5% adjusted operating profit to revenue ratio is just pathetic. Current pricing dynamics (2026): OpenAI already undercuts on many tiers ( flagship output tokens significantly cheaper than equivalent Claude Opus). Nano/mini models offer 5–10x advantages for volume work. Anthropic holds edges in long-context flat pricing and certain reasoning quality. OpenAI after Helios Rack Ownership, At $0.0003–$0.0005/M effective costs, OpenAI gains massive headroom to: ~Aggressively discount high-volume agentic tiers or bundles. ~Offer “unlimited” enterprise plans or usage-based models that Anthropic struggles to match without margin erosion. ~Target cost-sensitive, high-throughput agent deployments (dev tools, automation platforms) where token bills explode. Enterprises facing $ millions in monthly agentic bills will migrate to the provider delivering better economics at scale. OpenAI’s combination of strong models (o-series reasoning) + lowest TCO positions it to erode Anthropic’s enterprise share, especially as agentic becomes the dominant token consumer. Cheaper tokens expand the total addressable market dramatically. This feeds the data/model improvement loop, justifying further capex. AMD benefits from proven scale pulling in more customers (Meta, Oracle, Microsfot, Amazon, Softbank, TensorWave, LumaAI ... already aligned on Helios). Conclusion: Dr. Lisa Su has been laser focused on inference economics since at least 2022–2023, repeatedly emphasizing that the real battleground for AI scalability would be TCO, power efficiency (TDP), and ultimately tokens per dollar and per watt not just raw training FLOPS. While many viewed inference as a secondary, commoditized workload, Dr. Su architected AMD’s roadmap around rack-scale systems optimized for high-volume, sustained inference that would dominate as models matured and usage exploded. Helios represents the culmination of that multi-year bet: a fully integrated, open platform designed precisely for the economics of massive token throughput. This deep, strategic partnership with OpenAI starting with the 1GW Helios deployment in H2 2026 and scaling to 6GW, is the embodiment of that shared vision. Both companies foresaw a future where agentic AI models evolve to become extraordinarily token-hungry: autonomous agents executing complex, iterative workflows with planning, tool use, verification loops, and long-context reasoning. These workloads can consume 100x+ more tokens per task than traditional chat or single-turn generation, driving exponential demand as capabilities improve and enterprises deploy them at scale. By owning and optimizing this massive Helios capacity at GW scale, OpenAI achieves inference costs as low as $0.0003–$0.0005 per million tokens. This structural cost advantage allows OpenAI to absorb the coming token explosion profitably, dramatically lower effective pricing for enterprises, and win high-volume agentic workloads from higher-cost competitors like Anthropic. What was once a prohibitive monthly token bill becomes an affordable accelerator for productivity and innovation. The OpenAI-AMD alliance validates Dr. Su’s prescient strategy and turns the Agentic flywheel into reality: Collapsing inference costs → explosive token consumption → richer data and better models → accelerate greater demand. This partnership doesn’t just address today’s economics, it positions both leaders at the center of the infrastructure buildout that will power AI’s next decade. By delivering the lowest inference economics at scale, OpenAI not only solves enterprise bill pain but gains a decisive weapon to win share from higher-cost rivals like Anthropic. And that is why OpenAI and $META will deploy EPYC Dense Rack Not Financial Advice! DYOR! Research Purpose Only!

Mike

84,951 görüntüleme • 3 ay önce

$MU $SNDK $LITE $VRT NVIDIA and Groq: 2nd and 3rd Order Strategic Infrastructure Effects and Market Implications Public reporting indicates NVIDIA has agreed to acquire Groq for approximately $20,000,000,000 in cash, while excluding Groq’s nascent cloud business from the transaction perimeter. The reported carve-out materially constrains the immediate, direct linkage from the acquisition to incremental, NVIDIA-controlled data center capacity build-out because GroqCloud appears to be the principal channel through which Groq hardware is currently monetized at scale as a service. The infrastructure-market implications therefore depend primarily on post-close product strategy: whether NVIDIA (1) commercializes Groq silicon as a distinct inference product line and drives broad deployment through OEM/ODM channels and partners, (2) uses the acquisition mainly to absorb IP and talent while de-emphasizing standalone Groq hardware volumes, or (3) uses Groq technology to reshape NVIDIA’s own inference systems and networking roadmaps. The dominant transmission mechanism into memory, networking, and facility infrastructure markets is the degree to which NVIDIA shifts incremental inference deployments away from GPU architectures that are tightly coupled to external high-bandwidth memory (HBM) and toward Groq’s current architecture, which emphasizes large on-chip SRAM, deterministic compiler-scheduled execution, and direct chip-to-chip connectivity. Independent and company-published materials describe Groq’s current-generation approach as having no external memory, keeping weights and KV cache on-chip during processing, and requiring model sharding across multiple chips due to limited on-chip SRAM per device. That architectural choice is directionally HBM-negative on a per-accelerator basis and ambiguous for DRAM, NAND, networking, power, and cooling on a per-token basis because the design can reduce memory wall losses and tail-latency overhead while potentially increasing the number of chips and interconnect endpoints required to serve large models and long-context workloads. HBM implications are the most mechanically straightforward but should be framed as second-derivative rather than absolute. If Groq-class inference silicon meaningfully displaces NVIDIA GPU-based inference deployments, incremental HBM bit demand tied to inference growth could be reduced relative to a GPU-only baseline because Groq’s current approach does not appear to attach HBM stacks to each accelerator. However, current market structure suggests HBM remains supply-constrained and is being pulled by multiple vectors including continued GPU training scale and high-capacity inference configurations, with leading suppliers signaling tight conditions extending beyond 2026. In that environment, reduced inference-driven HBM intensity could primarily reallocate scarce HBM supply toward higher-end training and premium inference GPUs rather than creating an outright volume collapse, preserving high utilization of HBM capacity while potentially affecting the slope of pricing power and capacity expansion urgency over a multi-year horizon. The key downside scenario for the HBM complex would be a durable architectural bifurcation where “good-enough” inference shifts disproportionately to HBM-less ASICs across a broad swath of deployments (latency-sensitive, batch-1, cost-per-token optimized), while training remains GPU-HBM dominated; such a split would reduce the portion of future inference compute that naturally monetizes through HBM content and could compress the incremental HBM-per-AI-dollar ratio. The key upside/neutral scenario for HBM is that the supply chain remains fully allocated regardless, with NVIDIA using any “freed” HBM to ship more high-end GPUs into training and long-context inference, especially as roadmaps increase HBM per GPU, sustaining robust aggregate bit demand even if inference becomes more heterogeneous. Conventional DRAM implications split into 2 channels: (1) DRAM wafer capacity diversion into HBM and (2) DDR content per server in AI clusters. Supplier commentary indicates that AI-driven memory demand is supporting elevated DRAM markets more broadly, and HBM production is resource-intensive versus conventional DRAM, tightening supply for DDR products in parallel. A meaningful NVIDIA pivot to an inference architecture that reduces HBM dependence could, at the margin, ease the most acute HBM-driven bottlenecks and allow memory manufacturers more flexibility in balancing DRAM mix, which could be modestly DDR-positive on the supply side (less crowding-out) even if it is DDR-neutral or slightly negative on the demand side (if per-node CPU/DDR requirements decline due to more efficient accelerator utilization). The dominant practical outcome is likely that DDR demand remains supported by broad AI server proliferation and increasing memory footprints at the system level (CPUs, networking stacks, caching layers, retrieval-augmented pipelines), while HBM remains the premium profit pool; therefore, any HBM displacement that increases total server volumes could indirectly keep DDR demand resilient even if DDR per accelerator is not rising materially. NAND flash implications are comparatively indirect and volume-driven rather than architecture-driven. Inference clusters require SSD capacity for model storage, container images, logging, and increasingly for fast local retrieval indices and embedding stores, but the storage footprint per unit of compute is typically smaller than in training pipelines that stage large datasets and checkpoints. If NVIDIA uses Groq to lower inference cost and latency enough to expand the total number of inference deployment locations (regional colocation, enterprise on-prem, sovereign footprints), aggregate SSD attach could rise through geographic fragmentation and replication of model artifacts across more sites, even if per-site storage is modest. The NAND effect is therefore likely to be demand-broadening and mix-positive (datacenter SSDs) but not a primary swing factor versus the macro AI capex cycle and consumer/device cycles. Hard disk drive (HDD) markets should see negligible direct sensitivity because nearline HDD demand is driven by bulk storage and cloud archiving economics, while inference acceleration choices primarily reshape compute and network layers; any HDD benefit would be a tertiary function of overall data center square footage expansion rather than a direct consequence of Groq silicon displacing GPUs. Optical networking implications require separating (1) intra-cluster back-end fabrics that connect accelerators and (2) front-end / data center interconnect (DCI) that connects sites and regions. Groq’s own positioning and third-party reporting suggest scaling beyond a single node or rack relies on high-bandwidth fabrics and, in some described configurations, optical interconnect scaling across hundreds of chips. If NVIDIA commercializes Groq at scale, 2 offsetting forces emerge: lower cost-per-token and improved latency could expand inference throughput and drive more east-west traffic, increasing demand for high-speed switching and optics; conversely, if Groq delivers materially higher utilization and tokens per unit of network bandwidth for certain workloads, the network required per served token could decline. Public NVIDIA materials already indicate an aggressive photonics roadmap aimed at scaling AI factories, including co-packaged optics (CPO) switches and explicit collaboration with Coherent and Lumentum in the silicon photonics supply chain. That linkage is important because it suggests that, independent of Groq, NVIDIA is already pushing optics integration deeper into the switch package to reduce power and increase resiliency; Groq increases the strategic incentive to reduce network power and latency if inference becomes even more distributed and latency-sensitive. For Lumentum and Coherent specifically, the net implication is less about “more optics versus fewer optics” and more about a shift in optics form factor and value capture. Co-packaged optics can reduce reliance on pluggable transceivers in some switch architectures while increasing demand for integrated photonic engines, lasers, fiber attach, packaging processes, and component-level supply. NVIDIA’s own announcements explicitly position Coherent and Lumentum as collaborators in creating the integrated silicon/optics process and supply chain for photonics switches. If Groq accelerates the transition to very large-scale fabrics (more endpoints, higher port speeds, tighter power envelopes), that tends to pull forward CPO adoption and amplifies demand for the underlying photonics components even if the conventional pluggable module TAM is structurally pressured over time. If Groq instead pushes inference toward smaller, more localized pods (closer to users, more regional colocation), that can be optics-positive for DCI and metro connectivity because more sites must be interconnected at high bandwidth with low latency, favoring coherent optics and high-speed interconnect between facilities. The principal risk for optics suppliers is timing and margin structure: a faster move to NVIDIA-driven integrated photonics could concentrate bargaining power and compress margins for commoditized transceiver modules while favoring suppliers with differentiated lasers, integration capability, and qualification depth in NVIDIA’s CPO ecosystem. AEC and copper interconnect implications hinge on whether Groq deployment increases the density of short-reach links inside racks and rows. High-speed copper remains structurally advantaged at very short distances on cost, power, and serviceability, but reaches become constrained as lane speeds and aggregate bandwidth rise, creating a role for active electrical cables (AECs), retimers, and signal-conditioning silicon. Credo explicitly positions its AEC products as enabling reliable lossless 800G connectivity for AI clusters, and the company has highlighted participation at NVIDIA GTC with content focused on extending PCIe/CXL using AECs, indicating relevance to next-generation system topologies that require longer reach and higher signal integrity than passive copper can deliver. If NVIDIA turns Groq into a widely deployed inference card or chassis product, the likely near-term effect is AEC-positive because (1) more inference throughput tends to increase top-of-rack connectivity requirements, (2) distributing inference across more racks and sites increases short-reach links per unit of delivered service, and (3) PCIe-attached accelerator architectures tend to require robust signal conditioning as systems move to PCIe 6.x and beyond. Groq workshop materials explicitly reference GroqCard and GroqNode form factors, reinforcing that PCIe-attached deployment has been central to Groq’s current packaging strategy. The main countervailing risk is that Groq’s deterministic chip-to-chip fabric could be implemented primarily through backplanes and direct board-level connectivity that reduces the need for merchant AECs inside the box; in that case, incremental AEC demand would concentrate more in rack-to-switch and node-to-fabric links rather than within-chassis chip fabrics. Astera Labs implications are connectivity-architecture sensitive and, on balance, skew positive if NVIDIA increases heterogeneity and disaggregation in AI systems. NVIDIA has publicly positioned NVLink Fusion as a pathway for partners to build semi-custom AI infrastructure and has explicitly identified Astera Labs as a partner in that ecosystem, with Astera describing NVLink-related solutions expanding its connectivity platform across PCIe, CXL, and Ethernet plus fleet observability software. A Groq acquisition increases the probability that NVIDIA offers a broader menu of accelerators (training GPUs, inference-focused ASICs) and therefore increases the importance of scalable, high-reliability connectivity, retiming, switching, and telemetry across mixed topologies. If Groq silicon remains PCIe-attached in many deployments, PCIe 6.x retimers/switches and active cable modules become more central, aligning with Astera’s core portfolio. If NVIDIA instead integrates Groq concepts into scale-up fabrics (NVLink-like domains) or uses Groq to expand into inference “appliances” that must be rapidly deployed in colocation environments, the need for standard-compliant, serviceable connectivity with strong RAS/telemetry increases, again aligning with Astera’s positioning. Power equipment and cooling implications for Vertiv and adjacent suppliers should be viewed through the lens of rack power density, cooling modality (air vs liquid), and site deployment model (hyperscale campuses vs distributed colocation/enterprise). Groq claims its LPU and rack designs are “air-cooled by design” and require no complex cooling and power infrastructure, and third-party reporting has described Groq’s approach as relying on parallelism across many lower-power units rather than extreme per-chip performance. If NVIDIA scales Groq as a mainstream inference platform, the mix of data center cooling spend could shift modestly away from the highest-density liquid-cooled racks toward more air-cooled or hybrid deployments, particularly for inference pods placed in existing facilities that cannot easily retrofit for very high rack heat flux. That would be a mix headwind for suppliers most levered exclusively to high-end liquid cooling attachments per rack, but it is not necessarily a volume headwind for Vertiv given the company’s broad exposure to both power and cooling infrastructure and the likelihood that total AI deployment locations expand. Vertiv’s own industry commentary emphasizes that AI racks require higher power-density UPS, batteries, power distribution equipment, and switchgear capable of handling rapid load transients, and that hybrid cooling systems will evolve across deployment environments. Those statements align with a world where inference growth increases the count of powered racks and raises the operational complexity of power delivery even if per-rack density is lower than the most extreme training clusters. The most material infrastructure impact may occur outside the rack and upstream of the data hall: grid interconnects, substations, transformers, switchgear, generators, and utility-scale generation additions. Recent regulatory actions in the U.S. highlight that projected data center demand is already driving large planned increases in electricity generation capacity, underscoring that power availability is a binding constraint. In that context, an inference architecture that lowers joules per token could reduce the power required per unit of inference delivered, but it can also accelerate demand by lowering cost and improving latency, increasing the total volume of inference served (a classic rebound effect). The net outcome is likely continued, elevated demand for power infrastructure even if efficiency improves, with the key swing factor being whether AI capex remains on a multi-year growth trajectory or enters a digestion phase. Other data center infrastructure implications include server/ODM mix, facility design standardization, and networking architecture choices. If NVIDIA positions Groq-based inference as a broadly distributable “standard server + accelerator” solution rather than as an integrated, liquid-cooled rack like GB200 NVL72, spend could shift toward more conventional air-cooled server designs, higher unit volumes of mainstream racks, and faster deployment in colocation footprints, increasing demand for modular power rooms, busways, and rapidly deployable cooling solutions. If NVIDIA instead integrates Groq into its “AI factory” paradigm, the primary effect is likely acceleration of dense back-end fabric build-outs and a faster push toward photonics switching, increasing demand for fiber plant, connectors, and integrated optics supply chains while potentially compressing the lifecycle of transitional architectures based on pluggable optics and mid-reach copper. NVIDIA’s stated roadmap toward co-packaged optics and silicon photonics switches is already oriented toward scaling to very large GPU counts; adding a high-end inference ASIC increases the strategic importance of power-efficient, low-latency fabrics because inference economics become increasingly sensitive to network overhead as compute cost declines. Across the covered segments, the most defensible base case is limited near-term dislocation and a medium-term increase in uncertainty around memory intensity per unit of inference growth. HBM faces the clearest relative risk from an HBM-less inference platform, but supply tightness and GPU training roadmaps reduce the probability of an absolute demand shock over the next 12–24 months. Optical, AEC/copper, and power/cooling are more likely to remain volume-supported because they scale with endpoint count, deployment fragmentation, and total data center footprint, and those tend to rise when inference becomes cheaper and more widely deployed. The highest-conviction second-order effect is a shift in infrastructure mix: incrementally more distributed inference deployments (favoring colocation power/cooling standardization, DCI optics, and serviceable short-reach interconnect) and a gradual migration from pluggable optics toward integrated photonics in back-end fabrics (favoring suppliers positioned in the CPO ecosystem).

TheValueist

76,250 görüntüleme • 8 ay önce