Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M4 Max. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M4 Max Metal |
31.6 tok/s TG128 |
222 tok/s PP512 |
| 22,738.8 MiB |
| fabian |
| View Benchmark |
Qwen3.8-27B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max BLAS + Metal | 22.9 tok/s TG128 | 243 tok/s PP512 | 16,800.7 MiB | fabian | View Benchmark |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 583.3 tok/s TG128 @ 1 ctx | 9,998 tok/s PP512 | 1,371.1 MiB | basecompute | View Benchmark |
Qwen3-0.6B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 455.8 tok/s TG128 @ 1 ctx | 9,749 tok/s PP512 | 1,643.9 MiB | basecompute | View Benchmark |
gemma-3-1b-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 361.2 tok/s TG128 @ 1 ctx | 7,448 tok/s PP512 | 937.8 MiB | basecompute | View Benchmark |
gemma-3-1b-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 281.0 tok/s TG128 @ 1 ctx | 7,527 tok/s PP512 | 1,365.2 MiB | basecompute | View Benchmark |
Llama-3.2-1B-Instruct-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 371.7 tok/s TG128 @ 1 ctx | 5,963 tok/s PP512 | 1,089.1 MiB | basecompute | View Benchmark |
Qwen3-1.7B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 318.4 tok/s TG128 @ 1 ctx | 4,146 tok/s PP512 | 2,083.9 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 301.7 tok/s TG128 @ 1 ctx | 1,718 tok/s PP512 | 2,065.5 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Base-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 295.1 tok/s TG128 @ 1 ctx | 1,713 tok/s PP512 | 2,065.5 MiB | basecompute | View Benchmark |
Llama-3.2-1B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 320.5 tok/s TG128 @ 1 ctx | 6,058 tok/s PP512 | 1,673.3 MiB | basecompute | View Benchmark |
Qwen3-1.7B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 224.1 tok/s TG128 @ 1 ctx | 4,025 tok/s PP512 | 2,938.9 MiB | basecompute | View Benchmark |
Qwen3-4B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 157.5 tok/s TG128 @ 1 ctx | 1,725 tok/s PP512 | 3,897.3 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 205.4 tok/s TG128 @ 1 ctx | 1,721 tok/s PP512 | 2,950.5 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Base-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 208.0 tok/s TG128 @ 1 ctx | 1,727 tok/s PP512 | 2,950.7 MiB | basecompute | View Benchmark |
Llama-3.2-3B-Instruct-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 178.8 tok/s TG128 @ 1 ctx | 2,202 tok/s PP512 | 3,087.5 MiB | basecompute | View Benchmark |
Qwen3-4B-Instruct-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 121.9 tok/s TG128 @ 1 ctx | 1,711 tok/s PP512 | 3,897.2 MiB | basecompute | View Benchmark |
Qwen3-4B-Thinking-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 122.7 tok/s TG128 @ 1 ctx | 1,721 tok/s PP512 | 3,897.1 MiB | basecompute | View Benchmark |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 175.3 tok/s TG128 @ 1 ctx | 7,670 tok/s PP512 | 1,585.0 MiB | basecompute | View Benchmark |
Llama-3.2-3B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 133.1 tok/s TG128 @ 1 ctx | 2,218 tok/s PP512 | 4,770.8 MiB | basecompute | View Benchmark |
Qwen3-4B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 105.4 tok/s TG128 @ 1 ctx | 1,712 tok/s PP512 | 6,035.9 MiB | basecompute | View Benchmark |
Qwen3-4B-Instruct-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 104.3 tok/s TG128 @ 1 ctx | 1,723 tok/s PP512 | 6,035.8 MiB | basecompute | View Benchmark |
Qwen3-4B-Thinking-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 104.3 tok/s TG128 @ 1 ctx | 1,714 tok/s PP512 | 6,036.0 MiB | basecompute | View Benchmark |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 110.3 tok/s TG128 @ 1 ctx | 947 tok/s PP512 | 6,114.6 MiB | basecompute | View Benchmark |
gemma-4-E4B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 109.3 tok/s TG128 @ 1 ctx | 2,552 tok/s PP512 | 3,326.8 MiB | basecompute | View Benchmark |