๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Vamsi Batchu
@vamsibatchuk
Member of technical staff @GoogleDeepMind - design nerd - currently building @GoogleAIStudio // My views
๊ฐ€์ž… October 2011
2.5K ํŒ”๋กœ์ž‰ ์ค‘    9.3K ํŒฌ
Introducing Gemini 3.8 Flash TTS and Flash-Lite TTS from ๐Ÿ—ฃ๏ธ๐ŸŽ™๏ธ . Our most expressive audio generation models yet !!! here are some of the awesome capabilities from these models.. ๐Ÿ† #1# Hume Voice Design (71.4) and accent modeling (60.8) ๐ŸŽจ Design unique voices from scratch with natural language prompts ๐ŸŒ 100+ languages and dialects, including regional accents ๐Ÿ“š 2,000+ production-ready voices in the library โฑ๏ธ Voice replication from a 30-second sample with consent checks ๐ŸŽฌ Line-by-line direction for pacing, emotion, laughs, and pauses ๐Ÿ‘ฅ Native two-speaker scene staging from one script ๐Ÿ“– Hours of long-form audio with minimal speaker drift & much more...
๋” ๋ณด๊ธฐ