for september, we’re cutting the price of Dedicated Inference on H100s from $5.49/hr to $3.99/hr
new + existing deployments get the lower price automatically
deploy gemma 4, qwen3/3.5, gpt-oss, llama, nemotron 3.5 lightning models, or bring your own lora for a fine-tuned model
try it today:
顯示更多