🚀 MiniMax H3, accelerated on Day 1 with Sol Engine!
Within just 4.5 hours, our agent-native Sol Video Inference Engine achieved:
⚡ 3.95× end-to-end speedup over Diffusers
⚡ 2.80× speedup over SGLang
🎬 8× NVIDIA GB200, 1344×768, 24 FPS, 124 frames
The acceleration combines kernel fusion and graph capture, cross-step caching, and training-free sparse attention powered by Sol-Attn—with no distillation, fine-tuning, LoRA, or offline calibration.
We are especially excited to see a powerful open-weight model like MiniMax H3 released to the community. Open models are essential for pushing video generation research, systems optimization, and real-world deployment forward.
We hope this is the beginning of a much more vibrant open-source video generation ecosystem—and Sol Engine will keep working to make the latest models faster and easier to deploy from day one.
🔗