註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Andi Marafioti
@andimarafioti
leading multimodal research @huggingface (prev @unity)
加入 April 2022
698 正在關注    8.5K 粉絲
Speech-to-speech no longer needs speech-to-text! Until now, our stack was VAD -> STT -> LLM -> TTS. Now it can send audio directly to multimodal LLMs: VAD → MLLM → TTS No STT. The model understands your voice. Now go build better voice agents!
顯示更多
0
37
1.4K
149
轉發到社區