vLLM Blog· vLLM Team, Inferact, Red Hat, and NVIDIA·· 1 天前AI 评分62
vLLM 支持 NVIDIA Vera Rubin NVL72,AgentX 吞吐达 GB200 NVL72 的 7.8 倍
vLLM Support for NVIDIA Vera Rubin NVL72: 7.8x Throughput over GB200 NVL72
AI 导读
vLLM 宣布已支持 NVIDIA Vera Rubin NVL72,提供每日容器构建,并可在该平台运行 DeepSeek、Kimi、GLM 和 MiniMax 等模型。
来源:vLLM Blog · vllm.ai