Register and share your invite link to earn from video plays and referrals.

Andi Marafioti
@andimarafioti
leading multimodal research @huggingface (prev @unity)
Joined April 2022
698 Following    8.5K Followers
Speech-to-speech no longer needs speech-to-text! Until now, our stack was VAD -> STT -> LLM -> TTS. Now it can send audio directly to multimodal LLMs: VAD → MLLM → TTS No STT. The model understands your voice. Now go build better voice agents!
Show more
0
37
1.4K
149
Forward to community