Five first-place results in MLPerf® Inference v6.1. 603,023 tokens/sec on DeepSeek R1 at 72 GPUs and preview-category results on the Nebius @nvidia Vera Rubin NVL72. We were one of only two submitters with results on that hardware.
Full results:
#MLPerf# #MLCommons# #Inference#