Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M4 Pro. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M4 Pro Metal |
44.5 tok/s TG128 @ 1 ctx |
1,081 tok/s PP512 |
| 5,525.7 MiB |
| basecompute |
| View Benchmark |
Qwen3-8B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 31.6 tok/s TG128 @ 1 ctx | 392 tok/s PP512 | 10,365.8 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 79.3 tok/s TG128 @ 1 ctx | 732 tok/s PP512 | 7,880.1 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 73.7 tok/s TG128 @ 1 ctx | 731 tok/s PP512 | 8,231.5 MiB | basecompute | View Benchmark |
gpt-oss-20b-MXFP4 BaseRTmxfp4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 72.7 tok/s TG128 @ 1 ctx | 734 tok/s PP512 | 10,012.8 MiB | basecompute | View Benchmark |
gemma-3-1b-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 224.9 tok/s TG128 @ 1 ctx | 3,430 tok/s PP512 | 940.8 MiB | basecompute | View Benchmark |
Qwen3-0.6B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 284.7 tok/s TG128 @ 1 ctx | 4,787 tok/s PP512 | 1,645.8 MiB | basecompute | View Benchmark |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 208.6 tok/s TG128 @ 1 ctx | 3,571 tok/s PP512 | 1,366.9 MiB | basecompute | View Benchmark |
Gemma-4 26B-A4B IT (smart Q4_0, QAT-lossless) llama.cppQ4_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro BLAS + Metal | 74.7 tok/s TG128 | 757 tok/s PP512 | 14,426.1 MiB | invocation | View Benchmark |