I’ve been spending some time with two DGX Sparks recently, with DeepSeek V4 Flash, a 284-billion-parameter model, being my go-to model of choice running across both of them. They’re connected over a ConnectX-7 cable with 128 GB of unified memory each, and can push anywhere from 30 to 50 tokens per second most of the time.