가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Alok
@analogalok
Mechatronics Engineer AI belongs on your device. • Offline inference • No subscriptions. Teaching you to own your AI Intelligence Stack
가입 November 2013
223 팔로잉 중    4.6K 팬
Ziphu achieved 3× performance gains and Nvidia level cost parity on domestic Chinese chips for serving GLM 3.5 Flash!. The wildest detail in the GLM-5.3-Flash release: Ziphu used a GLM 5.3 infrastructure agent to write custom GPU kernels, debug bottlenecks, and build an inference engine on top of SGLang, creating a self optimizing feedback loop that achieved 3× performance gains and Nvidia level cost parity on domestic Chinese chips.
더 보기