Register and share your invite link to earn from video plays and referrals.

Vaibhav (VB) Srivastav
@reach_vb
founder mode @OpenAI | ex @huggingface | F1 fan | Here for @at_sofdog’s wisdom | *opinions my own
Joined June 2017
297 Following    55.7K Followers
Codex analysed production traffic, improved load balancing, rewrote production GPU kernels and ran hundreds of experiments on its own speculative-decoding model. The kernel improvements reduced end-to-end serving costs by 20%, while speculative decoding improved token-generation efficiency by more than 15%.
Show more