Compare
Loading community benchmarks…
Loading community benchmarks…
Put two chips, runtimes, or runtime versions side by side. Everything else is held fixed or called out, so a ratio is worth exactly what the matched runs say.
Ratios read side A relative to side B. Geometric mean over 2 like-for-like pairs.
| Facet | Apple M5 Pro | Apple M1 Max |
|---|---|---|
| Model | Qwen3-1.7B | Qwen3-1.7B-Q4, Qwen3-1.7B-Q8 |
| Runtime version | BaseRT 0.2.4 | BaseRT 0.2.6 |
| Decode workload | TG128 | TG128 @ 1 ctx |
Each row is a configuration present on both sides. Values are per-cell medians; ratios read Apple M5 Pro relative to Apple M1 Max.
| Configuration | Decode A | Decode B | Ratio | Prefill A | Prefill B | Ratio | Runs A / B |
|---|---|---|---|---|---|---|---|
Qwen3-1.7B BaseRTQ4 | 222.9 | 219.7 | 1.01× | 7,654 | 2,244 | 3.41× | 2 / 1 |
Qwen3-1.7B BaseRTQ8 | 132.4 | 153.9 | 0.86× | 7,655 | 2,223 | 3.44× | 2 / 1 |
Top 4 of 4 runs by decode throughput.
| Model / format | Chip / backend | Decode | Prefill | Contributor | Date | Report |
|---|---|---|---|---|---|---|
basecompute/Qwen3-1.7B BaseRT 0.2.4Q4 | Apple M5 Pro Metal | 235.4 TG128 | 8,040 PP512 | arki05 | View Benchmark | |
basecompute/Qwen3-1.7B BaseRT 0.2.4Q4 | Apple M5 Pro Metal | 210.3 TG128 | 7,268 PP512 | arki05 | View Benchmark | |
basecompute/Qwen3-1.7B BaseRT 0.2.4Q8 | Apple M5 Pro Metal | 139.7 TG128 | 8,043 PP512 | arki05 | View Benchmark | |
basecompute/Qwen3-1.7B BaseRT 0.2.4Q8 | Apple M5 Pro Metal | 125.1 TG128 | 7,267 PP512 | arki05 | View Benchmark |
Top 2 of 2 runs by decode throughput.
| Model / format | Chip / backend | Decode | Prefill | Contributor | Date | Report |
|---|---|---|---|---|---|---|
Qwen3-1.7B-Q4 BaseRT 0.2.6Q4 | Apple M1 Max Metal | 219.7 TG128 @ 1 ctx | 2,244 PP512 | basecompute | View Benchmark | |
Qwen3-1.7B-Q8 BaseRT 0.2.6Q8 | Apple M1 Max Metal | 153.9 TG128 @ 1 ctx | 2,223 PP512 | basecompute | View Benchmark |