가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Mia
@MiaAI_lab
Building with AI & LLMs | Insights, recipes, tools & honest experiments.
가입 July 2022
413 팔로잉 중    36.1K 팬
If you already have 4x DGX Sparks, you might not missing out much waiting for the 512 GB M5 Ultra. Here's why: Memory capacity: both 512 GB unified. Peak memory bandwidth: 4x Sparks: ~1092 GB/s aggregate (TP=4) M5 Ultra: ~1,200 GB/s They're almost identical on what matters most for local LLMs - memory capacity and bandwidth. Where the DGX Sparks still pull ahead Prefill is generally stronger, concurrency scales more cleanly, and the entire CUDA ecosystem (vLLM, TensorRT-LLM, etc.) is simply more mature for multi-node inference right now. Video and image generation will also probably be noticeably faster on the sparks due to CUDA. CUDA is still king. The Mac is cleaner and more efficient as a single box. But if you already own the 4x DGX Sparks, you can continue sleeping well.
더 보기