@MichaelGannotti: Doing the work on 2 @NVIDIAAI DGX Sparks
Summary
DeepSeek-V4.1-Flash AI model was used on NVIDIA DGX Spark units to automatically perform kernel work, read project documents, and open a pull request, showcasing real AI automation in software development.
View Cached Full Text
Cached at: 09/15/26, 05:47 PM
Doing the work on 2 @NVIDIAAI DGX Sparks
Tom Turney (@no_stp_on_snek): DeepSeek-V4.1-Flash on 2 Sparks. thinking high, TP=2. 55M tokens. Actual kernel work, not a demo.
it read the notes and prds, ran A against B, checked a win, then opened a PR.
i watched it publish. i did not write it.
The fix at a glance: causal attention is a “triangle”. the
Similar Articles
@MichaelGannotti: With model training done Nemo has brought @deepseek_ai V4 Flash back online on my 2 node @NVIDIAAI DGX Cluster and I ca…
Michael Gannotti shares that he has brought DeepSeek's V4 Flash AI model back online on his NVIDIA DGX cluster after completing model training, allowing him to resume local inference for running agents.
dgx sparks and new models my tests and results
This article presents test results for AI models like DeepSeek V4 Flash and Qwen3.8 on NVIDIA DGX Sparks hardware, detailing performance metrics, context lengths, and benchmark scores with operational insights.
@no_stp_on_snek: It's still pretty awesome what a single spark can do. Thanks @NVIDIAAI
A user shares that they replicated running DeepSeek-V4-Flash-0731 on a DGX Spark using antirez's DwarfStar-4 setup, confirming impressive performance on a single device.
@MichaelGannotti: https://x.com/MichaelGannotti/status/2074486390432149979
A full-stack evaluation of NVIDIA's Nemotron-3 Mamba-Transformer hybrid models on DGX Spark hardware, including architecture analysis, quantization (NVFP4), benchmark results, and introduction of the open-source smf-bench testing suite.
@ViC305: I DID IT!! DeepSeek-V4-Flash-Vision EXL3 MixedK is now running VISION + DSpark speculative decoding together on ONE DGX…
User @ViC305 successfully runs DeepSeek-V4-Flash-Vision with EXL3 MixedK and DSpark speculative decoding on a single DGX Spark, achieving improved performance and fixing technical issues for multimodal AI deployment.