So why AMD and Intel can't do this???
They didn't predict that LLMs are so good that can write efficient kernels on NUMA chips with ease?
No MTP, No PD disaggregation, Pure TP, still beats NVIDIA's Vera Rubin NVL72 on a third-party model, with A0 stepping. And B0 is 25% better.
NVIDIA GPUs become HBM wrappers.