Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Our self-improving agents optimized the full @vLLM_project inference stack, with up to 16% more throughput and interactivity for DeepSeek v4 Pro and Z.AI GLM 5.2 on B200s (no MTP). Every change was verified and our agents got better and faster at it with each iteration.

40,023 görüntüleme • 4 gün önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar