Skip to content
View original post on X: vLLMOfficial· 46/100AI score46/100

vLLM adds NVIDIA Vera Rubin support, reaching 7.8x GB200 throughput on MiniMax M3

AISummary

vLLM now supports NVIDIA Vera Rubin, and early results show more than 7.8x the throughput of GB200 running MiniMax M3 on AgentX. The post is the first in a five-part thread on the Rubin bring-up, which involved Inferact, NVIDIA, Red Hat AI, and the vLLM community.

Post on XView on X
vLLMVerified on X
@vllm_project

vLLM now supports NVIDIA Vera Rubin. The early results show more than 7.8x the throughput of GB200 on MiniMax M3 on AgentX.

@inferact, @NVIDIAAI, @RedHat_AI, and the vLLM community have been bringing vLLM up on Rubin since it was announced. Here is where things stand.

🧵 1/5

Source: vLLM · x.comPublished · added here