Register and share your invite link to earn from video plays and referrals.

Inferact
@inferact
Building the future of inference through @vllm_project
Joined December 2025
5 Following    7K Followers
Thanks for the shoutout @SemiAnalysis_ ! Full breakdown linked here:
ALERT ALERT ALERT 🚨 🚨 🚨 VLLM MAINTAINERS HAVE JUST SHOWN THAT TPUv7 CAN GET 700 tok/s/user,  56% BETTER PERFORMANCE THAN NVIDIA GB200 NVL72 THROUGH MEGAKERNEL OPTIMIZATION ON KIMI K3. As we said awhile ago, the TPU externalization of software is full steam ahead. This is ultra important to follow the progress of this.
Show more