Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M5 Pro. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M5 Pro Metal |
64.7 tok/s TG128 |
3,414 tok/s PP512 |
| 6,047.3 MiB |
| arki05 |
| View Benchmark |
basecompute/Qwen3.5-2B-Base BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 218.2 tok/s TG128 | 2,645 tok/s PP512 | 2,070.5 MiB | arki05 | View Benchmark |
basecompute/Qwen3.5-2B-Base BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 131.4 tok/s TG128 | 2,642 tok/s PP512 | 2,955.1 MiB | arki05 | View Benchmark |
basecompute/gemma-4-26B-A4B-it BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 51.4 tok/s TG128 | 3,253 tok/s PP512 | 17,830.8 MiB | arki05 | View Benchmark |
basecompute/Qwen3.6-27B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 10.6 tok/s TG128 | 355 tok/s PP512 | 18,211.3 MiB | arki05 | View Benchmark |
basecompute/Qwen3.8-27B BaseRTQ4· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 10.6 tok/s TG128 | 355 tok/s PP512 | 21,918.2 MiB | arki05 | View Benchmark |
basecompute/Qwen3-30B-A3B-Thinking-2507 BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 65.1 tok/s TG128 | 3,692 tok/s PP512 | 1,930.1 MiB | arki05 | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 66.0 tok/s TG128 | 1,748 tok/s PP512 | 1,118.6 MiB | arki05 | View Benchmark |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 356.2 tok/s TG128 | 14,552 tok/s PP512 | 2,333.7 MiB | arki05 | View Benchmark |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 358.3 tok/s TG128 | 14,509 tok/s PP512 | 2,333.1 MiB | arki05 | View Benchmark |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 358.1 tok/s TG128 | 14,344 tok/s PP512 | 2,342.6 MiB | arki05 | View Benchmark |
Qwen3-0.6B llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 283.0 tok/s TG128 | 14,942 tok/s PP512 | 2,565.6 MiB | arki05 | View Benchmark |
Llama-3.2-3B-Instruct llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 117.7 tok/s TG128 | 3,318 tok/s PP512 | 3,904.2 MiB | arki05 | View Benchmark |
Llama-3.2-3B-Instruct llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 115.2 tok/s TG128 | 3,315 tok/s PP512 | 3,940.1 MiB | arki05 | View Benchmark |
Llama 3.2 3B Instruct llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 76.6 tok/s TG128 | 3,516 tok/s PP512 | 5,237.3 MiB | arki05 | View Benchmark |
Qwen3-4B-Instruct-2507 llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 94.4 tok/s TG128 | 2,568 tok/s PP512 | 4,860.1 MiB | arki05 | View Benchmark |
Qwen3-4B-Instruct-2507 llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 92.6 tok/s TG128 | 2,566 tok/s PP512 | 4,902.8 MiB | arki05 | View Benchmark |
Qwen3-4B-Instruct-2507 llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 60.8 tok/s TG128 | 2,686 tok/s PP512 | 6,558.2 MiB | arki05 | View Benchmark |
Gemma-4-E4B-It llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 72.1 tok/s TG128 | 2,263 tok/s PP512 | 5,321.4 MiB | arki05 | View Benchmark |
Gemma-4-E4B-It llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 69.4 tok/s TG128 | 2,235 tok/s PP512 | 5,463.1 MiB | arki05 | View Benchmark |
Gemma-4-E4B-It llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 47.9 tok/s TG128 | 2,346 tok/s PP512 | 8,387.1 MiB | arki05 | View Benchmark |
Qwen3-8B llama.cppQ2_K The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 72.8 tok/s TG128 | 1,477 tok/s PP512 | 5,611.1 MiB | arki05 | View Benchmark |
Qwen3-8B llama.cppQ2_K The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 69.6 tok/s TG128 | 1,465 tok/s PP512 | 5,825.1 MiB | arki05 | View Benchmark |
Qwen3-8B llama.cppQ3_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 61.5 tok/s TG128 | 1,418 tok/s PP512 | 6,415.0 MiB | arki05 | View Benchmark |
Qwen3-8B llama.cppQ3_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 60.9 tok/s TG128 | 1,414 tok/s PP512 | 6,588.2 MiB | arki05 | View Benchmark |