On a DGX Spark, what you run still decides how fast it goes.
For example, just turning on MTP in llama.cpp you can make Qwen3.8 27B on a single Spark far faster. Lots more cases like this everyday.
We'll push this box to the limits until a new version is out...