Qwen Image 2.1 is here! 🖼️
A 7B params native image generation and editing model, with up to 10 image references
The model comes with it's own prompt enhancement LLMs, integrated with diffusers 🧨 and ComfyUI
▶️ on Spaces
hum anything, get a finished song 🎤
Hum-to-Song is an YuE2 LoRA that turns humming into full songs 🎼
hum in, a produced track, verse, chorus, and the tune you hummed coming back around
▶️ on Spaces
Meridian by @ViggleAI just dropped on @huggingface
an H3 based video model that allows you change the camera angle and timing of any existing video or image 🎥↔️
freeze it in time, or reimagine the shot
🏀 try on spaces
LTX Ripple is a new efficient IC LoRA approach for video editing ✏️🎞️
edit just the first frame and have the edits ripple into the rest of the video. lands perfect edits, incredibly fast! LTX 2.5 based
▶️
Breeze TTS 2 by @BreezeBlueX dropped on Hugging Face as the #1# open weights TTS model on @ArtificialAnlys
It does
🎨 Voice Design (prompt a voice)
🎙️ Voice Reference (upload a voice)
🎛️ Voice Direction (upload & prompt a voice)
vibe it on Spaces
▶️
SenseNova-U1.5-8B-MoT is HERE, it's a big deal! 🔥 open weights matching Nano Banana 2 on benchs 🍌
no VAE, no text encoder, no DiT. a mixture of transformers: text and image tokens use different weights, attend to each other, denoise in pixel space
▶️
text-to-motion that can follow precise commands and sequences
"a person walks forward, then sits down on the floor" — and it does them in that order 🚶🪑
PRISM (1.4B) hands you back SMPL-X params, so you can drop it on your own rig
▶️ on Spaces
Tencent just dropped SCoPE on Hugging Face 🧊
the next step of camera control for video models 🎥
feed a single image input + draw a 3D camera path and get a perfect shot
which of your photos would you fly a camera through?
▶️ on Spaces
find & replace, but for a voice recording 🎙️
FireRedTTS3 is out on @huggingface, an omni-TTS model that can do voice design, voice cloning and a cool new feature: speech in-painting - or voice editing!
▶️ on Spaces
extremely precise control just landed to Krea-2 as a LoRA! 🎨
krea2-turbo-bbox adds precise layout-controlled generation to Krea 2: drag the boxes, regenerate, the composition obeys, text boxes included
cooked by jimmycarter 🧑🍳
▶️ on Spaces
Google published gmn, a parametric differentiable 3D head model that runs on CPU 🔥
Every facial movement and expression, differentiable and malleable 🤪
Interactive demo on @huggingface Spaces
▶️
NVIDIA’s ARDY is a glimpse of where AI animation is heading: real time, open source and you can play with it right now on @huggingface Spaces
Type what the character should do and generate a motion sequence in seconds. Give a sentence a body.