Compare
Loading community benchmarks…
Loading community benchmarks…
Put two chips, runtimes, or runtime versions side by side. Everything else is held fixed or called out, so a ratio is worth exactly what the matched runs say.
Ratios read side A relative to side B. Geometric mean over 1 like-for-like pair.
| Facet | BaseRT | llama.cpp |
|---|---|---|
| Model | Qwen3.6-27B, Qwen3.6-27B-cuda-q4mix, Qwen3.6-27B-cuda-q8, Qwen3.6-27B-Q4, Qwen3.6-27B-Q8 | Qwen3.6-27B |
| Chip | Apple M1 Max, Apple M3 Ultra, Apple M4 Max, Apple M5 Max, Apple M5 Pro, NVIDIA GB10 | AMD Radeon RX 7900 XT, AMD Radeon RX 7900 XT (RADV NAVI31), Apple M5 Pro, Tesla T10/Tesla T10/Tesla T10/Tesla T10 |
| Backend | CUDA, Metal | BLAS + Metal, CUDA, ROCm, Vulkan |
| Conditioning | warmup_only | runtime_native_warmup |
| Decode workload | TG128, TG128 @ 1 ctx | TG128 |
| Harness schema | basert-benchmark-harness/1 | computearena-measurements/1 |
Each row is a configuration present on both sides. Values are per-cell medians; ratios read BaseRT relative to llama.cpp.
| Configuration | Decode A | Decode B | Ratio | Prefill A | Prefill B | Ratio | Runs A / B |
|---|---|---|---|---|---|---|---|
Qwen3.6-27B Apple M5 Pro | 13.3 | 14.9 | 0.90× | 360 | 375 | 0.96× | 2 / 4 |
Top 13 of 13 runs by decode throughput.
| Model / format | Chip / backend | Decode | Prefill | Contributor | Date | Report |
|---|---|---|---|---|---|---|
Qwen3.6-27B-Q4 BaseRT 0.2.6Q4 | Apple M3 Ultra Metal | 36.8 TG128 @ 1 ctx | 281 PP512 | basecompute | View Benchmark | |
basecompute/Qwen3.6-27B BaseRT 0.2.4Q4 | Apple M5 Max Metal | 31.3 TG128 | 603 PP512 | basecompute | View Benchmark | |
Qwen3.6-27B-Q4 BaseRT 0.2.6Q4 | Apple M4 Max Metal | 30.3 TG128 @ 1 ctx | 221 PP512 | basecompute | View Benchmark | |
Qwen3.6-27B-Q8 BaseRT 0.2.6Q8 | Apple M3 Ultra Metal | 22.1 TG128 @ 1 ctx | 280 PP512 | basecompute | View Benchmark | |
Qwen3.6-27B-Q4 BaseRT 0.2.6Q4 | Apple M1 Max Metal | 17.9 TG128 @ 1 ctx | 80 PP512 | basecompute | View Benchmark | |
basecompute/Qwen3.6-27B BaseRT 0.2.4Q4 | Apple M5 Pro Metal | 16.1 TG128 | 364 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B-Q4 BaseRT 0.2.6Q4 | NVIDIA GB10 CUDA | 13.0 TG128 @ 1 ctx | 344 PP512 | basecompute | View Benchmark | |
Qwen/Qwen3.6-27B BaseRT 0.2.4Q4 | NVIDIA GB10 CUDA | 12.4 TG128 | 1,140 PP512 | basecompute | View Benchmark | |
Qwen3.6-27B-cuda-q4mix BaseRT 0.2.6Q4 | NVIDIA GB10 CUDA | 12.3 TG128 @ 1 ctx | 1,127 PP512 | basecompute | View Benchmark | |
Qwen3.6-27B-Q8 BaseRT 0.2.6Q8 | Apple M1 Max Metal | 11.0 TG128 @ 1 ctx | 83 PP512 | basecompute | View Benchmark | |
basecompute/Qwen3.6-27B BaseRT 0.2.4Q8 | Apple M5 Pro Metal | 10.6 TG128 | 355 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B-cuda-q8 BaseRT 0.2.6Q4 | NVIDIA GB10 CUDA | 8.4 TG128 @ 1 ctx | 1,146 PP512 | basecompute | View Benchmark | |
Qwen3.6-27B-cuda-q8 BaseRT 0.2.6Q4 | NVIDIA GB10 CUDA | 7.5 TG128 @ 1 ctx | 1,141 PP512 | basecompute | View Benchmark |
Top 11 of 11 runs by decode throughput.
| Model / format | Chip / backend | Decode | Prefill | Contributor | Date | Report |
|---|---|---|---|---|---|---|
Qwen3.6-27B llama.cpp b1 (e64c0ea)Q4_0 | Tesla T10/Tesla T10/Tesla T10/Tesla T10 CUDA | 48.0 TG128 | 1,174 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q3_K_M | AMD Radeon RX 7900 XT (RADV NAVI31) Vulkan | 36.8 TG128 | 739 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q4_K_M | AMD Radeon RX 7900 XT (RADV NAVI31) Vulkan | 33.7 TG128 | 782 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q4_K_M | AMD Radeon RX 7900 XT (RADV NAVI31) Vulkan | 32.5 TG128 | 783 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q4_K_M | AMD Radeon RX 7900 XT ROCm | 29.8 TG128 | 841 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q3_K_M | AMD Radeon RX 7900 XT ROCm | 29.7 TG128 | 807 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q4_K_M | AMD Radeon RX 7900 XT ROCm | 29.3 TG128 | 840 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q3_K_M | Apple M5 Pro BLAS + Metal | 16.7 TG128 | 373 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q4_K_M | Apple M5 Pro BLAS + Metal | 15.2 TG128 | 376 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q4_K_M | Apple M5 Pro BLAS + Metal | 14.5 TG128 | 374 PP512 | arki05 | View Benchmark | |
Qwen3.6-27B llama.cpp b10809 (5266f24da)Q6_K | Apple M5 Pro BLAS + Metal | 11.9 TG128 | 377 PP512 | arki05 | View Benchmark |