註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Ahmad
@TheAhmadOsman
Founder & CEO @OsmanticAI — Accelerating Opensource & Self-hosted / Local AI Adoption • I moderate GPUs on r/LocalLLaMA
加入 February 2011
465 正在關注    78.4K 粉絲
The optimizations in MLX are horrid, DHH is right Apple doesn’t have a hardware problem they have a software problem
Faster chips, new bottleneck 🤯 Two clustered M5 Ultra Macs ran SLOWER than one: 35 tok/s vs 53. Which is strange, considering my experiments with M3 Ultra were the opposite. Keep the GPUs awake and they jump to 58 tok/s.
顯示更多
0
29
259
8
轉發到社區