注册并分享邀请链接,可获得视频播放与邀请奖励。

Rohan Siva
@_rsiva
research @decagonai ml & robotics research @utaustin
加入 August 2025
29 正在关注    46 粉丝
multimodal inference feels pretty underexplored, especially when you move beyond standard transformer architectures. fun to share what we learned when scaling our text to speech model to 10x throughput
显示更多
Fast text serving ≠ fast speech. Here's how we got our time to first audio under 30ms with nearly 10x more audio throughput.