← All guides

Is the DGX Spark Worth It? $4,699 Price vs Real Speed

The DGX Spark is the most argued-about box in local AI. It packs a GB10 Grace Blackwell chip and 128GB of unified memory into something that sits on a desk — and its 273 GB/s of memory bandwidth is a sixth of an RTX 5090's. Whether it is worth $4,699 depends entirely on which of three buyers you are, and most people asking are the wrong one.

Bottom Line

  • Price: $4,699 NVIDIA list price. The cheapest Amazon US unit was $4,999.99 in September 2026, and the NVIDIA store was out of stock. NVIDIA added an official $700 in February, blaming memory supply. Correction (2026-09-13): the ASUS Ascent GX10 no longer undercuts it. OEM GB10 boxes rose to $5,400-8,200 by September 2026 (every GB10 price).
  • Speed: stack-dependent to a degree nobody expects: gpt-oss 120B decode ran ~11.7 tok/s on Ollama, ~38.5 on tuned llama.cpp, ~50 on SGLang — same box, ~4x spread.
  • The real product: CUDA parity with your deployment target, on a desk. Not tokens per second.
  • Verdict: worth it for NVIDIA-stack developers. Not worth it as a pure inference box — a ~$3,500 Ryzen AI Max+ 395 machine does that job for about $1,200 less.

DGX Spark price: what $4,699 actually buys

GB10 Grace Blackwell, 128GB unified memory, ConnectX networking for pairing two units, and the full NVIDIA software stack in a quiet desktop box. The price history matters: it launched at $3,999, and in February 2026 NVIDIA raised it to $4,699 in an official notice citing memory supply constraints — the same shortage that reset RAM and GPU prices everywhere.

DGX Spark price history

DatePriceWhat changed
October 2025$3,999Launch price, Founders Edition
February 2026$4,699Official NVIDIA increase of $700 for memory supply
September 2026$4,999.99Cheapest Amazon US unit; NVIDIA store out of stock

The list price did not change after February. The street price did. In September 2026 you pay about $300 over list, or 25% over the launch price. For the OEM boxes on the same chip, see every GB10 box price.

The spec that defines it is 273 GB/s of memory bandwidth. An RTX 5090 moves ~1,792 GB/s — about 6.5x more. Token generation is bandwidth-bound, so no amount of Blackwell compute makes single-stream chat fast on this box. That is not a flaw discovered by reviewers; it is the design tradeoff: capacity and ecosystem over bandwidth.

The benchmark chaos, explained

The Spark’s launch reviews disagreed with each other by 4x, and both sides were measuring honestly:

Stackgpt-oss 120B decodeSource
Ollama~11.7 tok/scommunity review benchmarks
llama.cpp (tuned)~38.5 tok/sposted by the llama.cpp author
SGLang~50 tok/sNVIDIA developer forum testing

Two lessons. First, on this machine the software stack is worth more than 3x the hardware delta between it and its rivals — if you buy one, do not run Ollama defaults and conclude you were robbed. Second, all three numbers obey the same ceiling: gpt-oss 120B is MoE with a small active set, which is why it moves at all. A dense 70B reading ~40GB per token caps in single digits — if dense 70B is your goal, that money buys a much better rig.

The three buyers it is right for

  1. The deploy-to-NVIDIA developer. You prototype locally, ship to A100/H100/Blackwell in production, and need identical CUDA/TensorRT behavior. The Spark is the only 128GB desk box that gives you that. This is its honest, load-bearing use case.
  2. The DGX/NIM ecosystem team. Desk-side node, same containers, same tooling as the fleet.
  3. The two-Spark buyer. NVIDIA’s 200B-class and 1M-context configurations pair two units over ConnectX. Niche, real, and nothing else at this price does it.

If you are none of these, you are buying a badge.

What to buy instead

Same chip, different price: the ASUS Ascent GX10 — the same GB10 with 128GB — listed at $3,999 (1TB) in August 2026, but the ASUS US store listed it at $6,999 on 2026-09-13. Check every GB10 box price before you pick one. On that date the GX10 cost $2,300 more than the Spark’s $4,699 list price.

Same job, less money: a 128GB Ryzen AI Max+ 395 box ($3,499.99 GMKtec EVO-X2 or $3,799 Minisforum MS-S1 MAX on the vendor stores, 2026-09-17; it was ~$2,000-2,400 in August, sold out) is bandwidth-limited to the same order and runs the same MoE models at broadly comparable speeds — our model picks for it show gpt-oss 120B at 31-55 tok/s. No CUDA. If your stack is llama.cpp/Ollama anyway, that is $900-1,200 saved.

More speed, less capacity: if your models fit in 24-48GB, a used-GPU build demolishes every unified-memory box on generation speed.

The three boxes from this page:

Amazon affiliate links — we earn a small commission at no cost to you.

Sources

See Also

Need OpenClaw fixed live?

Remote rescue sessions for gateway, auth, tunnel, VPS, and model access problems.

See Rescue Session

Read next

Two Used RTX 3090s or One RTX 5090? 48GB Slow vs 32GB Fast
Dual used RTX 3090s cost $2,000-2,600 for 48GB of VRAM. One RTX 5090 costs $4,300-5,000 for 32GB. The 2026 price spike flipped this comparison: the dual build is now half the price AND holds a 70B. Here is the honest tradeoff, including the 700W problem.
Which Ryzen AI Max+ 395 Mini PC Should You Buy?
GMKtec EVO-X2 vs Framework Desktop vs Beelink GTR9 Pro vs Minisforum MS-S1 MAX. Same APU, same ~256 GB/s. As of September 2026 the 128GB boxes cost $3,449 to $4,349, and the old $1,999 price is gone. What to actually buy in 2026.
DGX Spark vs Strix Halo for Local LLMs: $4,699 vs $3,500 128GB Boxes
Compare NVIDIA DGX Spark and AMD Strix Halo (Ryzen AI Max+ 395) 128GB mini PCs for local LLMs: bandwidth, CUDA vs ROCm, MoE performance, and pricing.
Best Models to Run on AMD Ryzen AI Max+ 395 Boxes
Best local LLMs for AMD Ryzen AI Max+ 395 (Strix Halo) 128GB mini-PCs in 2026. Qwen3-30B-A3B at ~100 tok/s, gpt-oss 120B at 31-55 tok/s, Llama 4 Scout at ~18 tok/s, dense 70B at ~5 tok/s. Framework Desktop, GMKtec EVO-X2, HP Z2 Mini G1a compared against DGX Spark and Mac Studio — with 2026 prices, which the memory shortage has moved a long way.