註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Albert Gu
@_albertgu
assistant prof @mldcmu. chief scientist @cartesia. leading the ssm revolution.
加入 December 2018
78 正在關注    21.8K 粉絲
Within the span of a week, we launched streaming TTS (text-to-speech) and STT (speech-to-text) models that topped the leaderboards. I'm incredibly proud of the research team for their relentless pursuit of improvement, which have unlocked new state-of-the-art audio models on the Pareto frontier of speed and quality. As a research problem, speech requires fusing both text and audio and is the gateway to general multimodal models. We built Sonic-3.5 and Ink-2 from the ground up, developing multiple innovations along the way in a direction that will scale to general real-time intelligence. I've personally been deeply involved in building these models and more; it's been a blast working with the incredibly talented research team here @cartesia, and I can't wait to show the world what's coming next :)
顯示更多
0
3
181
16
轉發到社區