注册并分享邀请链接,可获得视频播放与邀请奖励。

Cyrus
@cyrusasg
research lead @decagonai prev @harvard post-training, inference, mlsys
加入 March 2022
255 正在关注    552 粉丝
multimodal inference doesn't fit the mold of standard text autoregressive generation. text-to-speech models feature different architectures, different states, and different batch shapes. we rebuilt our tts serving around that and simultaneously reduced our time to first audio while improving throughput by several fold.
显示更多