Builds / $20k tier

The Spark cluster

Three DGX Sparks, 384GB pooled over ConnectX: the compact rack that punches up.

TECHNICAL DIFFICULTY: 2 OF 4 ยท SEAT AND GO
How hard is this to assemble? three boxes to cable together, but each is plug in and go

NVIDIA's answer to the Mac Studio: 128GB of coherent memory per box, and the NVLink-style interconnect to pool them. Three Sparks cost less than one 512GB Ultra and bring CUDA-native inference, which matters for the newest model drops that land on NVIDIA stacks first. The trade: noisier, hotter, and per-dollar memory is worse than Apple.

WHO THIS BUILD IS FOR

You want CUDA, you want 384GB of pooled memory, and you want it in three boxes that slide under a desk instead of a rack. Developers testing against the newest NVIDIA-first model drops, teams pooling one budget, anyone whose lab smells faintly of ambition.

ESTIMATED TOTAL
$14,097
at MSRP, not yet price-checked
first check: next Monday

The parts

NVIDIA DGX Spark 128GB
NVIDIA DGX Spark 128GB x3 128GB 273GB/s nvidia
128GB each, pooled with the built-in ConnectX networking. 384GB total, NVFP4 native.
at MSRP, not yet price-checked
$14,097

What it runs

GLM-5.3-Flash
4-bit is 160GB of weights in 384GB pooled: a real fit with context room. Its home-turf CUDA stack.
DeepSeek V4.1 Flash
3-bit is about 207GB in 384GB pooled: fits with room to spare.
Qwen3.8-Omni-Flash
4-bit is about 62GB of weights in 384GB pooled: fits easily. Voice, vision, and text in one model.
WHAT IT WILL NOT RUN

The 2.8T class (Kimi K3 is 700GB at 2-bit). And the pooled interconnect is fast but not magic: tensor parallelism across three boxes adds latency you will notice in single-user chat, less in batch.

POWER, NOISE, AND THE ROOM IT LIVES IN

About 720W for the trio at load, and they are datacenter-adjacent loud. A closet with a door, or a garage shelf. Ethernet and power are the whole install.

WANT MORE? THE $30k TIER UNLOCKS

The question that keeps showing up on X: I have 20 to 30 thousand dollars, what do I buy? One answer is a single 512GB box that runs everything. The other is a rack of NVIDIA that trades silence for CUDA-native speed. This tier is where those two answers live, priced and compared.

See the $30k tier →