
NVIDIA AI Infrastructure
@NVIDIAAIInfra • 77,077 subscribers
AI factories for the era of AI reasoning.
Videos

The NVIDIA Vera Rubin NVL72 compute tray. 200 AI petaFLOPs. Assembled in one minute. That's what a single-wide, third-generation NVIDIA MGX rack enables. No cables. No hoses. No fans. It's 100% liquid cooled at 45°C and brings together Vera Rubin Superchips, ConnectX-9 SuperNICs, and BlueField-4 DPUs to deliver the lowest token cost and best performance per watt. #NVIDIAVeraRubin
NVIDIA AI Infrastructure334,083 次观看 • 1 个月前

💡 You’re either a token producer or consumer. Gavin Baker, Chief Investment Officer and Managing Partner at Atreides Management, LP, sets the stage for the next phase of tokens and how they will be an incredibly vital asset for businesses. 📺 Watch the full #NVIDIAGTC Live Pregame Show:
NVIDIA AI Infrastructure182,233 次观看 • 4 个月前

10x more tokens/second per megawatt as measured by CoreWeave on NVIDIA Vera Rubin platform. That performance reflects a bigger shift in AI infrastructure economics: building AI factories that efficiently turn compute into tokens, and tokens into revenue. It takes extreme co-design across the full stack, combined with an open ecosystem built to innovate and scale. Dion Harris connects the economics to the technology with 5 critical questions every AI infrastructure leader should be asking across compute, networking, storage, software and security. ➡️
NVIDIA AI Infrastructure23,393 次观看 • 15 天前

🤝 Alphabet and NVIDIA are expanding their decade-long partnership to advance agentic AI, robotics, drug discovery, and more. This involves deep co-engineering with integrated platforms, open-source frameworks, and managed services. ✅ Google Cloud is one of the first to bring the NVIDIA Blackwell platform to the cloud. ✅ Google Distributed Cloud using NVIDIA Confidential Computing on #NVIDIABlackwell for enterprises to run Google Gemini on-premises. ✅ Integration of the NVIDIA AI platform across Vertex AI, Cluster Director and Google Kubernetes Engine. ✅ Bringing the NVIDIA Nemotron family of open models to Vertex AI Model Garden. Watch the full video to learn more ➡️
NVIDIA AI Infrastructure268,567 次观看 • 7 个月前

📣 Get a first look at the NVIDIA Photonics co-packaged optics switch with Lambda. At NVIDIA GB300 NVL72 scale, the network doesn't just move data between GPUs — it determines how fast your cluster thinks. Co-packaged optics cut switch power, reduce failure points, and deliver more tokens per watt. Here's what that looks like in practice. ➡️
NVIDIA AI Infrastructure84,302 次观看 • 2 个月前

The NVIDIA Vera Rubin NVL72 scale-up fabric, built on the sixth-generation NVIDIA NVLink. The NVLink 6 Switch trays connect 72 Rubin GPUs in an all-to-all configuration. The NVLink spine—over 5,000 copper cables make up the vertical NVLink traverse. It securely moves more data than the entire bandwidth of the internet. The single-wide NVIDIA MGX rack design unlocks a cable-free, horizontal NVLink scale-up traverse, running directly on PCB for resiliency at scale. This is how NVIDIA Vera Rubin NVL72 works as one, designed to deliver the lowest token cost and 10x more performance per watt than the previous generation. #NVIDIAVeraRubin
NVIDIA AI Infrastructure38,483 次观看 • 1 个月前

"Compute is revenues. Performance per watt is your revenues. Choosing the wrong architecture just because the chips are cheaper doesn't translate, doesn't make sense." — Jensen Huang, CEO of NVIDIA In the age of AI factories, compute is revenue — every token produced is profitable — making delivered performance per watt, reliability, and the long lifetime of these systems the core financial levers, not just peak specs. Hear Jensen explain it at his #NVIDIAGTC Taipei keynote:
NVIDIA AI Infrastructure61,179 次观看 • 3 个月前

🎬 Higgsfield AI 🧩 is accelerating professional video production, enabling millions of creators to produce content at unprecedented speed and scale, powered by NVIDIA Blackwell. They utilized our accelerated computing platform for: ✅ 30% faster training on NVIDIA HGX B200 and B300 systems, including on Nebius ✅ Infrastructure reliability through NVIDIA NVSentinel and DCGM ✅ Advanced creative control and optimized performance with NVIDIA CUDA libraries and InfiniBand networking Learn more ➡️
NVIDIA AI Infrastructure60,045 次观看 • 4 个月前

🤝 Arcee.ai, an American AI lab, built its frontier open model family on NVIDIA Blackwell Ultra. Its flagship MoE model, Trinity-Large-Thinking, was RL post-trained using NVIDIA NeMo open libraries. Arcee optimized agentic model inference with NVIDIA Dynamo and vLLM, along with NVIDIA accelerated networking, to deliver low token cost, achieving: ✅ 3T+ tokens served on OpenRouter in its first two months ✅ $0.90 per 1M output tokens ✅ 2nd on PinchBench for open agentic models Explore Arcee's open models, powered by NVIDIA ➡️
NVIDIA AI Infrastructure27,377 次观看 • 1 个月前

Our internal AI factory serves 4 trillion tokens a month to NVIDIA employees, and demand is still growing 40% month over month. In Episode 2 of AI Factory Insider, we get into what it actually takes to run that at scale: the architecture, the real use cases, and what we learned along the way. Catch the full episode ➡️
NVIDIA AI Infrastructure20,714 次观看 • 1 个月前

AI that responds in under 100ms doesn't happen by accident. Together AI VP of Kernels Dan Fu shares how they use NVIDIA CUDA, TensorRT-LLM, Dynamo and Together ATLAS on NVIDIA Blackwell to power ultra-low-latency inference and long context code generation for Cursor. From building megakernels that fuse an entire model's forward pass into a single launch to enabling real-time voice agents, Together AI is optimizing latency at every layer of the stack.
NVIDIA AI Infrastructure20,712 次观看 • 2 个月前

The #NVIDIAVeraRubin platform consists of five rack-scale systems for AI agents with a supply chain spanning over 350 factory sites across 30 countries with millions square feet of factory floor space. Vera Rubin is in full production, co-designed with NVIDIA DSX and DSX MaxLPS to deliver the lowest token cost and maximum tokens per watt. Congratulations to our partners who have their engineering racks up and running: CoreWeave, Dell Technologies, Microsoft, and Oracle Cloud.
NVIDIA AI Infrastructure16,701 次观看 • 1 个月前

✨ The energy at #NVIDIAGTC Taipei at #COMPUTEX2026 was incredible. The show was packed with the community and ecosystem partners that make this all possible. Loved seeing all the accelerated computing technologies on the showroom floor! Thank you to all the attendees, sponsors, and partners. We'll see you next time. ✌️
NVIDIA AI Infrastructure23,222 次观看 • 2 个月前

NVIDIA Vera Rubin is ramping into full production to power agentic AI factories worldwide. With a supply chain twice as large as NVIDIA Grace Blackwell, Vera Rubin provides manufacturing, cloud, and infrastructure partners with a POD-scale platform to fuel the buildout of tomorrow’s AI factories. Read the press release → #NVIDIAGTC
NVIDIA AI Infrastructure18,278 次观看 • 2 个月前

"Agentic AI changes the role of the CPU. The CPU is now the conductor and the GPU is the orchestra". 🎼 NVIDIA Vera is the first CPU built for AI agents — purpose-built from the ground up for how AI works today. Faster. More efficient. Ready for what's next. 🔲 #NVIDIAGTC
NVIDIA AI Infrastructure15,062 次观看 • 3 个月前

Last week at #CES2026, our CEO Jensen Huang unveiled #NVIDIARubin, our new extreme-codesigned, six chip AI platform, alongside major advancements in open models and physical AI. On the show floor, we showcased some of the AI infrastructure and #acceleratedcomputing technologies that's powering the next era of AI. Thank you to everyone who stopped by the NVIDIA Showcase and we’ll see you next time. 👋
NVIDIA AI Infrastructure30,115 次观看 • 7 个月前

🌌 20 Terabytes of data. That's what the Rubin Observatory captures every night to map dark matter — but traditional processing takes months. By leveraging CPU/GPU-accelerated pipelines on multi-GPU, multi-node systems, NVIDIA cuPhoton accelerates image loading by 15,000x and processing by 8,400x, turning months of waiting into minutes. #ISC26
NVIDIA AI Infrastructure11,168 次观看 • 2 个月前






