NVIDIA DGX Spark

gpu4.pics
submitted
NVIDIA DGX Spark GB10 Grace-Blackwell
NVIDIA
running qwen3:14b
23.9
tokens / sec
Generation
23.9tok/s
higher is better
Prompt / prefill
tok/s
higher is better
Time to first token
lower is better
Composite score
Model size
Runtime
ollama
ollama
gpu4.pics/r/nvidia-dgx-spark-z14i

All benchmarks on this rig

ModelQuantGenPromptTTFTPowerPerf/W$/MtokScore
qwen3:14b23.9

How it stacks up — qwen3:14b

NVIDIA DGX Spark GB10 Grace-Blackwell
23.9
Apple M3 Max
35.3