๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

apolinario (poli)
@multimodalart
ML Engineer for Art and Creativity @HuggingFace
๊ฐ€์ž… July 2021
707 ํŒ”๋กœ์ž‰ ์ค‘    16.4K ํŒฌ
MiniMax H3 image+audio to video feed an audio into h3 instead of generating it. i was testing this flow, got shocked with how good the lipsync and movement is implemented it on the i2v version in diffusers, made a demo for it. what do you think? ๐Ÿ”Š โ–ถ๏ธ
๋” ๋ณด๊ธฐ