Show your rig. Prove your tokens.
A beautiful, open gallery of homelab & local-AI rigs — real tokens/sec, perf-per-watt, and cost, on the hardware people actually run. 70 rigs and counting.
▣ AMD Instinct MI50 32GB
32 GBgpt-oss 120B· Q8
58
tok/s
Apple M3 Max
128 GBIBM Granite 4.1 8B· Q4_K_M
57.6
tok/s
Apple M2 Ultra
192 GBDolphin Mixtral 8x7B· Q5
37
tok/s
Apple M3 Ultra
512 GBDeepSeek R1 671B· Q4
18
tok/s
▤ NVIDIA Tesla V100
32 GBQwen3 235B· Q4
15
tok/s
▤ NVIDIA L40S
44 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
355
tok/s
▤ NVIDIA RTX 3090 Ti
24 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
354
tok/s
▤ NVIDIA H100
80 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
335
tok/s
▤ NVIDIA RTX 3090
24 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
330
tok/s
▤ NVIDIA RTX A6000
48 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
315
tok/s
▤ NVIDIA A100 80GB
80 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
308
tok/s
▤ NVIDIA RTX 3090
24 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
308
tok/s
▤ NVIDIA Tesla V100
32 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
248
tok/s
▤ NVIDIA Tesla V100
32 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
248
tok/s
▤ NVIDIA RTX PRO 6000 Blackwell Workstation Edition
95 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
215
tok/s
▤ NVIDIA RTX 4090
24 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
207
tok/s
Apple M4 Max
128 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
184
tok/s
▤ NVIDIA RTX 5090
32 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
170
tok/s
Apple M4 Max
128 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
156
tok/s
▤ NVIDIA RTX 6000 Ada Generation
48 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
131
tok/s
▤ NVIDIA H100
80 GBMeta Llama 3.1 8B Instruct· Q4_K_M
120
tok/s
▤ NVIDIA A100 80GB
80 GBMeta Llama 3.1 8B Instruct· Q4_K_M
110
tok/s
▤ NVIDIA RTX 3090 Ti
24 GBMeta Llama 3.1 8B Instruct· Q4_K_M
110
tok/s
▤ NVIDIA RTX PRO 6000 Blackwell Server Edition
95 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
106
tok/s
▤ NVIDIA RTX PRO 6000 Blackwell Workstation Edition
95 GBMeta Llama 3.1 8B Instruct· Q4_K_M
103
tok/s
▤ NVIDIA RTX 3090
24 GBMeta Llama 3.1 8B Instruct· Q4_K_M
103
tok/s
▤ NVIDIA H200
141 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
98.5
tok/s
▤ NVIDIA RTX A6000
48 GBMeta Llama 3.1 8B Instruct· Q4_K_M
90.5
tok/s
▤ NVIDIA RTX 4090
24 GBMeta Llama 3.1 8B Instruct· Q4_K_M
90.3
tok/s
▤ NVIDIA H100
80 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
72.2
tok/s
▤ NVIDIA RTX PRO 6000 Blackwell Workstation Edition
95 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
70.8
tok/s
▤ NVIDIA A100 80GB
80 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
66.3
tok/s
▤ NVIDIA RTX 5090
32 GBMeta Llama 3.1 8B Instruct· Q4_K_M
66.3
tok/s
▤ NVIDIA RTX 3090 Ti
24 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
64.2
tok/s
▤ NVIDIA RTX 3090
24 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
59.9
tok/s
▤ NVIDIA RTX A6000
48 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
52
tok/s
Apple M4 Max
128 GBMeta Llama 3.1 8B Instruct· Q4_K_M
51.6
tok/s
▤ NVIDIA RTX 6000 Ada Generation
48 GBMeta Llama 3.1 8B Instruct· Q4_K_M
51.3
tok/s
▤ NVIDIA RTX 4090
24 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
49.1
tok/s
▤ NVIDIA RTX 5090
32 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
45.5
tok/s
Apple M3 Max
128 GBqwen3:14b
35.3
tok/s
▤ NVIDIA RTX 6000 Ada Generation
48 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
28.1
tok/s
▤ NVIDIA RTX PRO 6000 Blackwell Server Edition
95 GBQwen2.5 14B Instruct· Q4_K_M· 14.8B
17.9
tok/s
▣ AMD EPYC 7352 24-Core Processor (znver2)
126 GBLlama 3.2 1B Instruct· Q4_K_M· 1.5B
14.3
tok/s
▤ NVIDIA A100 80GB
80 GBGoliath 120B· GGUF8-BIT
12
tok/s
▤ NVIDIA Tesla P40
24 GBQwen3 32B· Q4
5.9
tok/s