$NVDA is turning MGX into a broader inference platform by adding d-Matrix Raptor for frontier models that need far more memory than SRAM can support.
Groq covers ultra-fast SRAM inference while Raptor brings 2.3TB of 3D-DRAM per rack through the same NVLink ecosystem.