We're introducing Provisioned Throughput: reserved inference capacity for frontier open models, with token-based pricing and a 99% uptime SLA.
Serverless simplicity, guaranteed capacity, up to 90% lower cost vs. Opus 4.8.
Get started with MiniMax M3 + GLM-5.2, read more ๐งต