注册并分享邀请链接,可获得视频播放与邀请奖励。

SemiAnalysis
@SemiAnalysis_
加入 January 2024
34 正在关注    168K 粉丝
Great blog from Kevin Lu at @vllm_project about quality control at the inference engine. Unfortunately, @AnushElangovan has not provided enough stable AMD QA clusters to vLLM, which leads to an order of magnitude worse software quality on AMD versus CUDA. Multiple AMD CI vLLM fleet-wide outages happen every month.
显示更多