quote: MiniMax M3 is now running on NVIDIA Vera Rubin NVL72 with vLLM!🙌
Thanks to @vllm_project, @inferact, @NVIDIAAI, | Hanami
quote: MiniMax M3 is now running on NVIDIA Vera Rubin NVL72 with vLLM!🙌
Thanks to @vllm_project, @inferact, @NVIDIAAI, @RedHat_AI for bringing MiniMax M3 to Rubin. https://x.com/vllm_project/status/2108736734309896625 | vLLM now supports NVIDIA Vera Rubin. The early results show more than 7.8x the throughput of GB200 on MiniMax M3 on AgentX.
@inferact, @NVIDIAAI, @RedHat_AI, and the vLLM community have been bringing vLLM up on Rubin since it was announced. Here is where things stand.
🧵 1/5