Register and share your invite link to earn from video plays and referrals.

dylan ツ
@demian_ai
growth @nebiustf @nebiusai // ex @Scaleway // from silicon to token, inference and anything in between. Views are my own - not financial advice
Joined January 2022
2.6K Following    29.4K Followers
okay enough stonk talk, time for GPU talk $NBIS we just ran our widest MLPerf Inference round yet and the full rack Nebius system with NVIDIA GB300 NVL72 took first on DeepSeek R1 in both server and offline (about 603k and 690k tokens a second). simple version of what happened: 1. we got the new chips (including Vera Rubin, the latest NVIDIA iron, and we were one of only two labs to show it in this round) 2. we racked them (full GB300 NVL72, plus the smaller boxes people actually buy). 3. we put them to the test under MLPerf, the public inference benchmark where somebody else holds the stopwatch result: the full rack took #1# on DeepSeek R1 for live and batch traffic, and pushed gpt-oss 120B past a million tokens a second. When we grew from 8 GPUs to 72 (almost 9x!), speed scaled almost 9x too. One fast GPU is easy, a rack that stays fast when you multiply it is the product. another day another milestone
Show more
Five first-place results in MLPerf® Inference v6.1. 603,023 tokens/sec on DeepSeek R1 at 72 GPUs and preview-category results on the Nebius @nvidia Vera Rubin NVL72. We were one of only two submitters with results on that hardware. Full results: #MLPerf# #MLCommons# #Inference#
Show more