Five first-place results in MLPerf® Inference v6.1. 603,023 tokens/sec on DeepSeek R1 at 72 GPUs and preview-category results on the Nebius
@nvidia Vera Rubin NVL72. We were one of only two submitters with results on that hardware.
Full results:
#
MLPerf# #
MLCommons# #
Inference#