Register and share your invite link to earn from video plays and referrals.

Enze Xie
@xieenze_jr
Staff Research Scientist at NVIDIA, doing GenAI, CS PhD from HKU MMLab, interned at NVIDIA.
316 Following    2.2K Followers
🚀 MiniMax H3, accelerated on Day 1 with Sol Engine! Within just 4.5 hours, our agent-native Sol Video Inference Engine achieved: ⚡ 3.95× end-to-end speedup over Diffusers ⚡ 2.80× speedup over SGLang 🎬 8× NVIDIA GB200, 1344×768, 24 FPS, 124 frames The acceleration combines kernel fusion and graph capture, cross-step caching, and training-free sparse attention powered by Sol-Attn—with no distillation, fine-tuning, LoRA, or offline calibration. We are especially excited to see a powerful open-weight model like MiniMax H3 released to the community. Open models are essential for pushing video generation research, systems optimization, and real-world deployment forward. We hope this is the beginning of a much more vibrant open-source video generation ecosystem—and Sol Engine will keep working to make the latest models faster and easier to deploy from day one. 🔗
Show more
🚀 Sol Video Inference Engine is here! An agent-native, training-free full-stack accelerator for video diffusion. It auto-tunes cache + sparse attn + token pruning + quant + kernel fusion for any model/hardware/config. >2× end-to-end speedup on 64B Cosmos3-Super, 22B LTX-2.3 and 2B SANA-Video — near-lossless VBench quality, minimal human effort. Practical acceleration for real video gen deployment. 📄 Paper: 🌐 Project: 💻 Code: Proud of the team! 🎉
Show more