注册并分享邀请链接,可获得视频播放与邀请奖励。

Baseten
@baseten
Inference is everything.
加入 March 2021
340 正在关注    10.2K 粉丝
We serve Qwen3-TTS on vLLM-Omni at $3 per 1M characters. That's 90% lower in cost than comparable closed-source TTS APIs. Our engineers optimized a single-replica serving stack to get there. Details on the optimized stack and cost per concurrent stream here.
显示更多