正在加载视频...

视频加载失败

Our self-improving agents optimized the full @vLLM_project inference stack, with up to 16% more throughput and interactivity for DeepSeek v4 Pro and Z.AI GLM 5.2 on B200s (no MTP). Every change was verified and our agents got better and faster at it with each iteration.

39,695 次观看 • 3 天前 •via X (Twitter)

0 条评论

暂无评论

原始帖子的评论将显示在这里

相关视频