註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Pareton
@Pareton_ai
Faster. Cheaper. Verified on your inference workload. Optimize your inference on Bittensor Subnet 10.
加入 June 2026
8 正在關注    442 粉絲
We just put Pareton's engine on an H200 and beat a B200 on Qwen3.8-27B-FP8. B200, stock vLLM: 106.2 tok/s H200, Pareton engine: 252.4 tok/s 2.4x on a cheaper GPU. Same prompt, same output length, batch 1. Third-party verified:
顯示更多
0
7
46
14
轉發到社區