Hot or not?
Summary
The author is adding active cooling to their DGX stack to manage heat under extended load and plans to share results related to comphy and DeepSeek flash.
Similar Articles
Found a way to cool the DGX
A user reports successfully using tap water to cool a DGX server while running the Qwen3.5-122b model at high GPU utilization, maintaining safe temperatures.
Deepseek V4 flash performance on DGX Spark
A Reddit user shares their experience running DeepSeek V4 Flash on a dual-ASUS GX10 DGX Spark setup, detailing performance metrics, configuration, and power consumption, with throughput benchmarks across various context lengths.
dgx sparks and new models my tests and results
This article presents test results for AI models like DeepSeek V4 Flash and Qwen3.8 on NVIDIA DGX Sparks hardware, detailing performance metrics, context lengths, and benchmark scores with operational insights.
@MichaelGannotti: https://x.com/MichaelGannotti/status/2084186867000279436
A 14.7-hour soak test of DeepSeek V4 Flash on an NVIDIA DGX Spark with 971 requests shows zero crashes or errors, but throughput declined 28% due to thermal throttling, while TTFT and speculative acceptance remained stable.
@RayFernando1337: Let him cook. This is a fun time to be owning 2 DGX Sparks RN
A Twitter user discusses owning two NVIDIA DGX Sparks systems, while another user reports improved token throughput with a model update to DFlash2-7, achieving 67 tokens per second.