dgx-spark

Tag

Cards List
#dgx-spark

@MichaelGannotti: https://x.com/MichaelGannotti/status/2076024719371841537

X AI KOLs Timeline · 2026-07-11 Cached

A detailed report on optimizing a production vLLM serving configuration on NVIDIA's DGX Spark, correcting flags that were costing 34% MTP acceptance after reviewing 90+ official NVIDIA documents and running a 69-scenario tool evaluation.

0 favorites 0 likes
#dgx-spark

@WescheNex1q: 16 people chatting with Qwen3.6-35B at once ONE DGX Spark. This is a real capture, not a mockup: every token you see re…

X AI KOLs Timeline · 2026-07-09 Cached

A real-time demo shows 16 concurrent users chatting with Qwen3.6-35B on a single DGX Spark, achieving peak 440 tok/s total and 105 tok/s per user using NVFP4 + MTP-3 on vLLM.

0 favorites 0 likes
#dgx-spark

@MichaelGannotti: https://x.com/MichaelGannotti/status/2074486390432149979

X AI KOLs Timeline · 2026-07-07 Cached

A full-stack evaluation of NVIDIA's Nemotron-3 Mamba-Transformer hybrid models on DGX Spark hardware, including architecture analysis, quantization (NVFP4), benchmark results, and introduction of the open-source smf-bench testing suite.

0 favorites 0 likes
#dgx-spark

@ivanfioravanti: DGX Spark Context Benchmark on Qwen3.6-35B-A3B-UD-Q8_K_XL llamacpp script released by Mia. It's fast! Time to test qual…

X AI KOLs Timeline · 2026-07-06 Cached

Benchmark results for Qwen3.6-35B-A3B-UD-Q8_K_XL on DGX Spark using llama.cpp script by Mia, showing fast token generation times across various context lengths.

0 favorites 0 likes
#dgx-spark

@Tech2Wild: Is Anyone Here Regretting or Regret it ?

X AI KOLs Following · 2026-07-03 Cached

Miro warns that most people will regret buying a Mac or DGX Spark for local LLMs, and recommends the RTX Pro 6000 for serious use.

0 favorites 0 likes
#dgx-spark

Follow-up: GLM-5.2 NVFP4 on four DGX Sparks — the MTP mystery is solved, and it's now ~24 tok/s at 128K context

Reddit r/LocalLLaMA · 2026-07-03

A bug in vLLM's speculative decoding configuration for GLM-5.2 NVFP4 on four DGX Sparks was fixed, resolving a performance tradeoff and achieving ~24 tok/s at 128K context with MTP4.

0 favorites 0 likes
#dgx-spark

@DeRonin_: My current local AI setup: - 2x DGX Spark linked (256gb) > GLM 5.2 @ 2bit, reasoning + agent loops - Mac Studio M3 Ultr…

X AI KOLs Following · 2026-06-30 Cached

A user describes their fully local AI stack using multiple hardware devices running Chinese models like GLM, Qwen, and Kimi, claiming 87% cost savings compared to frontier models like GPT-5.5 and Opus 4.8, while noting plans to self-host video generation.

0 favorites 0 likes
#dgx-spark

@SpaceTimeViking: Announcing Orinth 1.0 AEON ULTIMATE UNCENSORED! BF16 and Quantized in NVFP4 for the DGX Spark / Blackwell arch. Preserv…

X AI KOLs Timeline · 2026-06-27 Cached

Announcing Orinth 1.0 AEON ULTIMATE UNCENSORED, a model with BF16 and NVFP4 quantization for DGX Spark/Blackwell architecture, claiming 200-300% performance improvement with working DFlash.

0 favorites 0 likes
#dgx-spark

@MiaAI_lab: What’s the best model you can run on your @NVIDIAAI DGX Spark? 1× DGX Spark * ⁠Qwen 3.6 35b NVFP4 - 256k ctx, 110 tok/s…

X AI KOLs Timeline · 2026-06-27 Cached

A tweet detailing the best AI models to run on Nvidia's DGX Spark, including Qwen 3.6 and DeepSeek v4 Flash variants, with token speeds and context lengths for single and multi-unit setups.

0 favorites 0 likes
#dgx-spark

@mr_r0b0t: To all 6,016 of you who follow me: HUGE THANKS We're just getting started! Please feel free to stop by my GitHub for a …

X AI KOLs Following · 2026-06-27 Cached

A user thanks followers and promotes GitHub repos with SM121 optimized containers for running local LLMs on DGX Spark (GB10) systems.

0 favorites 0 likes
#dgx-spark

@RayFernando1337: https://x.com/RayFernando1337/status/2070621713952579990

X AI KOLs Following · 2026-06-26 Cached

A detailed analysis on whether to run AI models locally or via API, covering hardware options like RTX 5090, RTX PRO 6000, and DGX Spark, with emphasis on memory vs bandwidth trade-offs, cost considerations, and privacy needs.

0 favorites 0 likes
#dgx-spark

1 rtx pro 6000 or 2 dgx sparks

Reddit r/LocalLLaMA · 2026-06-26

A comparison between a single RTX Pro 6000 GPU and two DGX Spark systems for AI compute tasks.

0 favorites 0 likes
#dgx-spark

Got GLM-5.2 + MTP speculative decode running on 4× DGX Spark (GB10) — and the build piece the public recipe is missing

Reddit r/LocalLLaMA · 2026-06-24

The author successfully ran GLM-5.2 with MTP speculative decoding on a 4× DGX Spark (GB10) setup, revealing a missing component in the public build recipe.

0 favorites 0 likes
#dgx-spark

@Ex0byt: Update: the road to GLM-5.2: we're getting there, folks! non-quantized, non-pruned DeepSeek-v4-Flash. 11tok/s on a sing…

X AI KOLs Timeline · 2026-06-23 Cached

Update on running a non-quantized DeepSeek-v4-Flash model at 11 tok/s on a single DGX Spark using sglang inference and a custom mega-kernel, progressing towards GLM-5.2.

0 favorites 0 likes
#dgx-spark

@aijoey: for all my new dgx spark owners. https://github.com/joeynyc/spark-doctor…

X AI KOLs Timeline · 2026-06-23 Cached

Spark Doctor is an open-source diagnostic CLI for NVIDIA DGX Spark that collects system, GPU, memory, Docker, and recipe data, applies specific rules, and outputs the likely cause and next steps for common issues.

0 favorites 0 likes
#dgx-spark

GLM 5.2 on 4x Sparks reasonable?

Reddit r/LocalLLaMA · 2026-06-17

A user asks about the feasibility of running GLM-5.2 at 4-bit quantization on four Ascend GX10s or DGX Sparks, wondering about speed and memory for 100k context.

0 favorites 0 likes
#dgx-spark

@MiaAI_lab: 3rd @NVIDIAAI DGX Spark on the way

X AI KOLs Following · 2026-06-15 Cached

NVIDIA announces the third generation of its DGX Spark AI supercomputer.

0 favorites 0 likes
#dgx-spark

Strix Halo desktop trying to compete against DGX Spark

Reddit r/LocalLLaMA · 2026-06-14 Cached

AMD launches the Ryzen AI Halo Developer Platform, a $3,999 mini PC with 128GB unified memory and Windows 11 support, competing with Nvidia's DGX Spark for local AI workloads.

0 favorites 0 likes
#dgx-spark

Deepseek V4 flash performance on DGX Spark

Reddit r/LocalLLaMA · 2026-06-01

A Reddit user shares their experience running DeepSeek V4 Flash on a dual-ASUS GX10 DGX Spark setup, detailing performance metrics, configuration, and power consumption, with throughput benchmarks across various context lengths.

0 favorites 0 likes
#dgx-spark

Dell confirms XPS laptop with NVIDIA N1X at Computex ( basically a DGX Spark GB10 for consumers with Windows )

Reddit r/LocalLLaMA · 2026-05-31

Dell confirms a new XPS laptop featuring the NVIDIA N1X chip, essentially a consumer version of the DGX Spark GB10, to be announced at Computex.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback