๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Tencent Hy
@TencentHunyuan
Tencent's foundation model for text, image, video, and 3D generation.
๊ฐ€์ž… July 2024
8 ํŒ”๋กœ์ž‰ ์ค‘    53K ํŒฌ
๐Ÿš€ AuK is officially here. Nano banana๐ŸŒ for audio An open-source foundation model for unified speech generation and editing. Natural-language instructions + reference audio. One interface. Zero-shot TTS. Instruction-controlled generation. Content editing. Whisper-conversion. De-accent. Timbre/style/emotion edit. Speed/Pitch control. Enhancement, denoising, multi-speaker and music separation. Also releasing AuK-Flash: 4-step inference. ~4.5ร— faster under matched conditions. Code, weights, and demo are live. Try it and share your feedback. ๐Ÿค— Paper & upvote: โญ GitHub & star:
๋” ๋ณด๊ธฐ