注册并分享邀请链接,可获得视频播放与邀请奖励。

Han Xiao
@hxiao
VP, AI @Elastic prev: founder & ceo @JinaAI_
加入 April 2009
283 正在关注    21.2K 粉丝
Model-specific inference engines are basically what everyone autoresearches and hill-climbs the hell out of for TPS after each Qwen release. In the early days I worked with MLX too, but these days I just trust oMLX to do the job.
显示更多
Meet Husky: a Model-Specific Inference (MSI) engine up to 4.5× faster than Apple's MLX Woof, Underdog's Pareto frontier model, now runs up to 730 tokens/sec on a MacBook Finally local models are as fast & capable. Try it now in - your personal private AI
显示更多