註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Rohan Siva
@_rsiva
research @decagonai ml & robotics research @utaustin
加入 August 2025
29 正在關注    46 粉絲
multimodal inference feels pretty underexplored, especially when you move beyond standard transformer architectures. fun to share what we learned when scaling our text to speech model to 10x throughput
顯示更多
Fast text serving ≠ fast speech. Here's how we got our time to first audio under 30ms with nearly 10x more audio throughput.