Apple M3 Max MacBook Pro

gpu4.pics
submitted
Apple M3 Max
128 GB unified Apple
running qwen3:14b
35.3
tokens / sec
Generation
35.3tok/s
higher is better
Prompt / prefill
tok/s
higher is better
Time to first token
lower is better
Composite score
Model size
Runtime
ollama
ollama
gpu4.pics/r/apple-m3-max-macbook-pro-fan4

All benchmarks on this rig

ModelQuantGenPromptTTFTPowerPerf/W$/MtokScore
qwen3:14b35.3

How it stacks up — qwen3:14b

Apple M3 Max
35.3
NVIDIA DGX Spark GB10 Grace-Blackwell
23.9