Two DGX Sparks. Three OEMs. One 200 GbE fabric and a methodological detour that produced the most interesting result of the review. We benchmarked dual-Spark clusters from Dell, GIGABYTE, and HP across multiple models and workload shapes, and found the OEMs landed within a narrow band. The real finding was about pipeline parallelism beating NVIDIA’s tensor parallelism default by 2.20x on GPT-OSS-120B once concurrency scaled. Link in bio.
#NVIDIA #DGXSpark #AI #inference