🚀 MiniMax H3, accelerated on Day 1 with Sol Engine!
Within just 4.5 hours, our agent-native Sol Video Inference Engine achieved:
⚡ 3.95× end-to-end speedup over Diffusers
⚡ 2.80× speedup over SGLang
🎬 8× NVIDIA GB200, 1344×768, 24 FPS, 124 frames
The acceleration combines kernel fusion and graph capture, cross-step caching, and training-free sparse attention powered by Sol-Attn—with no distillation, fine-tuning, LoRA, or offline calibration.
We are especially excited to see a powerful open-weight model like MiniMax H3 released to the community. Open models are essential for pushing video generation research, systems optimization, and real-world deployment forward.
We hope this is the beginning of a much more vibrant open-source video generation ecosystem—and Sol Engine will keep working to make the latest models faster and easier to deploy from day one.
🔗
🚀 Sol Video Inference Engine is here!
An agent-native, training-free full-stack accelerator for video diffusion. It auto-tunes cache + sparse attn + token pruning + quant + kernel fusion for any model/hardware/config.
>2× end-to-end speedup on 64B Cosmos3-Super, 22B LTX-2.3 and 2B SANA-Video — near-lossless VBench quality, minimal human effort.
Practical acceleration for real video gen deployment.
📄 Paper:
🌐 Project:
💻 Code:
Proud of the team! 🎉