注册并分享邀请链接,可获得视频播放与邀请奖励。

Ahmad
@TheAhmadOsman
Founder & CEO @OsmanticAI — Accelerating Opensource & Self-hosted / Local AI Adoption • I moderate GPUs on r/LocalLLaMA
加入 February 2011
466 正在关注    78.7K 粉丝
The optimizations in MLX are horrid, DHH is right Apple doesn’t have a hardware problem they have a software problem
Faster chips, new bottleneck 🤯 Two clustered M5 Ultra Macs ran SLOWER than one: 35 tok/s vs 53. Which is strange, considering my experiments with M3 Ultra were the opposite. Keep the GPUs awake and they jump to 58 tok/s.
显示更多
0
29
259
8
转发到社区