Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

We are getting 2 new devices from NVIDIA and Microsoft: 1. The DGX Station, with GB300 superchip and up to 748GB of memory. 2. The RTX Spark laptop, with 1 petaflop of AI performance and 128GB of unified memory. The hardware here is pretty bonkers! Think about it this way:

15,307 görüntüleme • 3 ay önce •via X (Twitter)

13 Yorum

Santiago profil fotoğrafı
Santiago3 ay önce

For the first time ever, you will be able to run a trillion-parameter model on your personal hardware. They announced many more things, but there's something else I want to share: They are running a 5-day mini-program to help developers ship agentic AI running on NVIDIA-accelerated Azure AI Foundry. Here is the Builder's Arcade program: • Day 1 - Run your first AI workload • Day 2 - Deploy your model as a service • Day 3 - Wire your model into an agentic workflow • Day 4 - Make your model faster and cheaper • Day 5 - Focus on governance, monitoring, and observability Go through this. It will take a few hours, and you'll learn a TON! Link to the Builder's Arcade: Here is the link to the announcements: Thanks to the Microsoft and NVIDIA teams for partnering with me on this post.

The AI Therapist profil fotoğrafı
The AI Therapist3 ay önce

1 petaflop inference, 748GB training. Both companies announcing because custom silicon is coming. This is what it looks like when a duopoly gets spooked.

Manav Gupta profil fotoğrafı
Manav Gupta3 ay önce

trillion param on local hardware changes everything for voice agents. latency was always the killer. this might actually fix it

Mevil | AI + Manufacturing profil fotoğrafı
Mevil | AI + Manufacturing3 ay önce

a dgx station class box on the factory floor changes the math — no cloud bill, no latency, production data never leaves the building. we've held off on vision models partly over connectivity 🤔 wonder how they'll price these vs a year of cloud inference?

Rush profil fotoğrafı
Rush3 ay önce

A time will come when you’ll be able to run big ai models on your laptop itself

UnsoldBanana profil fotoğrafı
UnsoldBanana3 ay önce

I'm not worried about the hardware. I'm more worried about Microsoft executing the software well. ARM laptops have existed before.

Ferbin profil fotoğrafı
Ferbin3 ay önce

Running voice agents on edge hardware, latency matters more than size. What's the actual inference speed on consumer GPUs? Trillion-param means nothing if it's slow.

Daniel Birker profil fotoğrafı
Daniel Birker3 ay önce

I really want the RTX Spark. Could run Deepseek like a dream.

Blum profil fotoğrafı
Blum3 ay önce

more and more hardware built specifically for AI is appearing

Alfero Chingono profil fotoğrafı
Alfero Chingono3 ay önce

That RTX Spark laptop with 128GB unified memory? That's huge for local LLM fine-tuning without the cloud. We've been waiting for this kind of power.

Luffysan profil fotoğrafı
Luffysan3 ay önce

RTX Sparkの128GB統合メモリは素晴らしいが、1ペタFLOPの性能を維持するには冷却と電源供給が鍵だ。

GooGZ AI profil fotoğrafı
GooGZ AI3 ay önce

Awesome. Pretty keen to see the laptop pricing when it eventually comes out. Hopefully some 'budget' 32GB and 48GB type options are released also.

Hussain Fakhruddin profil fotoğrafı
Hussain Fakhruddin3 ay önce

A petaflop laptop with 128GB unified memory changes local model math completely.

Benzer Videolar