註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Han Xiao
@hxiao
VP, AI @Elastic prev: founder & ceo @JinaAI_
加入 April 2009
283 正在關注    21.2K 粉絲
Model-specific inference engines are basically what everyone autoresearches and hill-climbs the hell out of for TPS after each Qwen release. In the early days I worked with MLX too, but these days I just trust oMLX to do the job.
顯示更多
Meet Husky: a Model-Specific Inference (MSI) engine up to 4.5× faster than Apple's MLX Woof, Underdog's Pareto frontier model, now runs up to 730 tokens/sec on a MacBook Finally local models are as fast & capable. Try it now in - your personal private AI
顯示更多