Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M4 Pro. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M4 Pro Metal |
113.9 tok/s TG128 @ 1 ctx |
1,149 tok/s PP512 |
| 2,953.8 MiB |
| basecompute |
| View Benchmark |
Qwen3.5-2B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 113.5 tok/s TG128 @ 1 ctx | 1,138 tok/s PP512 | 2,954.1 MiB | basecompute | View Benchmark |
Qwen3-4B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 90.3 tok/s TG128 @ 1 ctx | 734 tok/s PP512 | 3,898.9 MiB | basecompute | View Benchmark |
Qwen3-1.7B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 125.3 tok/s TG128 @ 1 ctx | 1,822 tok/s PP512 | 2,940.3 MiB | basecompute | View Benchmark |
Llama-3.2-1B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 179.3 tok/s TG128 @ 1 ctx | 2,692 tok/s PP512 | 1,675.1 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Base-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 176.8 tok/s TG128 @ 1 ctx | 1,150 tok/s PP512 | 2,069.3 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 179.1 tok/s TG128 @ 1 ctx | 1,150 tok/s PP512 | 2,070.3 MiB | basecompute | View Benchmark |
Qwen3-1.7B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 193.1 tok/s TG128 @ 1 ctx | 1,828 tok/s PP512 | 2,085.5 MiB | basecompute | View Benchmark |
Llama-3.2-1B-Instruct-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 204.6 tok/s TG128 @ 1 ctx | 2,685 tok/s PP512 | 1,090.2 MiB | basecompute | View Benchmark |