dgx-spark

Tag

Cards List
#dgx-spark

What's going on with DGX Spark? Price up $2k in 1 week?

Reddit r/LocalLLaMA ↗ · yesterday

A buyer reports that NVIDIA DGX Spark units are hard to find and that prices appear to have jumped roughly $2,000 in a week, with local Micro Center stock vanishing overnight, raising questions about supply allocation.

0 favorites 0 likes
#dgx-spark

@no_stp_on_snek: Same two GB10s. GLM-5.3-Flash, two stacks. One streamed call each, temperature 0.2, 384 tokens, prompts around 3.7k tok…

X AI KOLs Timeline ↗ · yesterday Cached

A head-to-head benchmark of two community inference stacks running the 320B GLM-5.3-Flash on a pair of NVIDIA DGX Sparks, comparing Entrpi's EXL3 setup against TensorFold with DFlash2 drafters across prose, code, and prefill workloads.

0 favorites 0 likes
#dgx-spark

Do you need some extra memory on your DGX Spark?

Reddit r/LocalLLaMA ↗ · 2d ago

This repository helps DGX Spark users utilize a spare GPU to free up memory for better context or quantization quality by offloading the draft model via remote inference with vllm modifications.

0 favorites 0 likes
#dgx-spark

@alexocheema: The handbook I wish I had when I bought my first DGX Spark

X AI KOLs Timeline ↗ · 6d ago Cached

A tweet sharing a personal handbook guide for new users of NVIDIA's DGX Spark AI computing system.

0 favorites 0 likes
#dgx-spark

@TheAhmadOsman: Mac Studio M5 Ultra vs 2x DGX Spark for DeepSeek V4 Flash - DGX Sparks: 2.41x faster on prefill (compute-bound) - Mac S…

X AI KOLs Timeline ↗ · 2026-09-24 Cached

The article compares the inference performance of Apple's Mac Studio M5 Ultra with two NVIDIA DGX Spark units when running the DeepSeek V4 Flash model, showing DGX Sparks are faster in prefill while Mac Studio is slightly faster in generation.

0 favorites 0 likes
#dgx-spark

@KyleHelseth: Nvidia marketplace of DGX sparks is out of stock for the first time

X AI KOLs Following ↗ · 2026-09-20 Cached

Nvidia's DGX Spark is out of stock for the first time on the marketplace, with prices rising at retailers, indicating strong AI hardware demand and potential economic implications.

0 favorites 0 likes
#dgx-spark

Built this yesterday with Qwen3.8-Flash-Next (NVFP4, 262K context) on a single NVIDIA DGX Spark

Reddit r/LocalLLaMA ↗ · 2026-09-18

A developer built a project in 8 hours using the Qwen3.8-Flash-Next model on a single NVIDIA DGX Spark, generating around 10k lines of code and consuming 800k tokens.

0 favorites 0 likes
#dgx-spark

@heyshrutimishra: My AI agent saves me 30-35 hours a week. I run a SaaS, and a brand agency. I do NOT have a big team. I have Hermes runn…

X AI KOLs Following ↗ · 2026-09-14 Cached

The author shares how an AI agent named Hermes, running on DGX Spark, automates business tasks like content management and client inquiries, saving 30-35 hours weekly, and argues that the agent market is at a pivotal moment akin to early mobile internet.

0 favorites 0 likes
#dgx-spark

Talk me out of buying a 3rd Spark

Reddit r/LocalLLaMA ↗ · 2026-09-13

A user asks on a forum if a 2x DGX Spark cluster can run the DeepSeek 4.1 Flash model, or if more Sparks are required, seeking community insights.

0 favorites 0 likes
#dgx-spark

@ViC305: 18 HOURS LATER: DeepSeek-V4.1-Flash is now quantized to 4.75 bpw EXL3 for a 4× DGX Spark TP4 target. The weights are DO…

X AI KOLs Following ↗ · 2026-09-11 Cached

DeepSeek-V4.1-Flash has been quantized to 4.75 bpw EXL3 for deployment on 4× DGX Spark, optimizing memory usage and enabling efficient local inference with plans for validation and further optimization.

0 favorites 0 likes
#dgx-spark

@QuixiAI: Two fixes for the NVIDIA open kernel driver on DGX Spark - Freed GPU memory now returns to the OS when a process exits …

X AI KOLs Timeline ↗ · 2026-09-05

NVIDIA's open kernel driver for DGX Spark received two fixes that return freed GPU memory to the OS when a process exits and enable huge pages for GPU page faults on system memory, boosting first-touch bandwidth from 0.4 to 19.6 GiB/s.

0 favorites 0 likes
#dgx-spark

@sgl_project: We just added recipes for DeepSeek-V4-Flash-Vision & DeepSeek-V4-Flash-0731 on 2x DGX Spark. https://docs.sglang.io/coo…

X AI KOLs Timeline ↗ · 2026-09-03 Cached

SGLang has added deployment recipes for DeepSeek-V4-Flash-Vision and DeepSeek-V4-Flash-0731 models on 2x DGX Spark hardware, with support for various configurations and optimizations.

0 favorites 0 likes
#dgx-spark

@ViC305: I DID IT!! DeepSeek-V4-Flash-Vision EXL3 MixedK is now running VISION + DSpark speculative decoding together on ONE DGX…

X AI KOLs Timeline ↗ · 2026-09-03 Cached

User @ViC305 successfully runs DeepSeek-V4-Flash-Vision with EXL3 MixedK and DSpark speculative decoding on a single DGX Spark, achieving improved performance and fixing technical issues for multimodal AI deployment.

0 favorites 0 likes
#dgx-spark

@QuixiAI: Like this is a legit bug DGX spark should get latest cuda the *instant* it's released @nvidia

X AI KOLs Following ↗ · 2026-09-02 Cached

A bug has been identified in DGX Spark, with the author urging NVIDIA to ensure it receives the latest CUDA updates without delay.

0 favorites 0 likes
#dgx-spark

@QuixiAI: I finally got a dgx spark First thing I noticed - they make it hard to install latest cuda. I might need to install van…

X AI KOLs Timeline ↗ · 2026-09-02 Cached

User @QuixiAI reports difficulty installing the latest CUDA toolkit on their NVIDIA DGX Spark, suggesting that installing vanilla Ubuntu may be necessary to resolve the issue.

0 favorites 0 likes
#dgx-spark

DGX Spark about to jump in price? Asus Ascent GX10 jumped from $3999 to $5999 today...

Reddit r/LocalLLaMA ↗ · 2026-09-01

The Asus Ascent GX10 desktop AI supercomputer has experienced a substantial price increase, leading to speculation that the DGX Spark system might also see a price jump soon.

0 favorites 0 likes
#dgx-spark

Qwen3.8-Flash-Next NVFP4 2xDGX Spark config: 50t/s decode, 2,900t/s prefill

Reddit r/LocalLLaMA ↗ · 2026-08-30

A user shares their optimized configuration for deploying the Qwen3.8-Flash-Next model on dual DGX Spark hardware, achieving up to 50t/s decode and 2,900t/s prefill speeds with technical patches and setup details.

0 favorites 0 likes
#dgx-spark

@shantanugoel: The config I arrived at after 2 days of sweeping through a bunch of hyper parameters and patches

X AI KOLs Following ↗ · 2026-08-28 Cached

Shantanu Goel shared a configuration recipe for optimizing Qwen 3.8 Flash Next on a single DGX Spark, tested for practical tasks and plans to benchmark it further.

0 favorites 0 likes
#dgx-spark

@RayFernando1337: Wow and repo is live!

X AI KOLs Following ↗ · 2026-08-28 Cached

Announcement that the repository for GLM 5.3 FLASH model is live, with initial benchmarks showing 59 tokens per second on DGX SPARK hardware and promises of more updates.

0 favorites 0 likes
#dgx-spark

CNBC Television: Nvidia partner with Perplexity AI to run locally in DGX Spark.

Reddit r/LocalLLaMA ↗ · 2026-08-25 Cached

NVIDIA and Perplexity AI have partnered to bring AI processing to local desktops via the DGX Spark device, focusing on cost reduction and data privacy, with cloud offloading for advanced reasoning tasks.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback