註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Paolo Rosson
@redp314
Head of Applied AI at Dext. Building | Quantum physics PhD from Oxford | Former Italian blindfolded Rubik’s cube record holder
加入 May 2015
2.3K 正在關注    1.7K 粉絲
Got Meta's new Muse Glimmer 30B running on my MacBook (M3 Max, 96GG) and tested the serving options available so far. Fastest right now: Ollama's MLX engine (DFlash included) at ~29 tok/s. Tuned llama.cpp: ~21. Raw mlx-vlm: ~10, not optimized yet. Numbers below if you're setting it up 👇
顯示更多