MiniMax H3 image+audio to video
feed an audio into h3 instead of generating it. i was testing this flow, got shocked with how good the lipsync and movement is
implemented it on the i2v version in diffusers, made a demo for it. what do you think? 🔊
▶️
Text to video, image to video with multiple key frames, video continuation, and native audio with lipsync, effects, and ambience. Access all advanced capabilities of FLUX 3 through OpenRouter's Video API.