跳到正文
原文
vLLM Blog· vLLM Team, Inferact, Red Hat, and NVIDIA·· 1 天前AI 评分62

vLLM 支持 NVIDIA Vera Rubin NVL72,AgentX 吞吐达 GB200 NVL72 的 7.8 倍

vLLM Support for NVIDIA Vera Rubin NVL72: 7.8x Throughput over GB200 NVL72

AI 导读

vLLM 宣布已支持 NVIDIA Vera Rubin NVL72,提供每日容器构建,并可在该平台运行 DeepSeek、Kimi、GLM 和 MiniMax 等模型。

来源:vLLM Blog · vllm.ai