Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M4 Max. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M4 Max Metal |
73.8 tok/s TG128 @ 1 ctx |
1,679 tok/s PP512 |
| 3,955.6 MiB |
| basecompute |
| View Benchmark |
muse-glimmer-30B-kquant-dynamic BaseRTpassthrough_gguf The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 22.8 tok/s TG128 @ 1 ctx | 254 tok/s PP512 | 609.8 MiB | basecompute | View Benchmark |
Qwen3.6-35B-A3B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 139.7 tok/s TG128 @ 1 ctx | 886 tok/s PP512 | 2,747.8 MiB | basecompute | View Benchmark |
Qwen3.5-35B-A3B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 140.1 tok/s TG128 @ 1 ctx | 895 tok/s PP512 | 2,747.6 MiB | basecompute | View Benchmark |
muse-glimmer-30B-kquant-17gb BaseRTpassthrough_gguf The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 27.2 tok/s TG128 @ 1 ctx | 266 tok/s PP512 | 605.8 MiB | basecompute | View Benchmark |
NVIDIA-Nemotron-3-Nano-30B-A3B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 169.1 tok/s TG128 @ 1 ctx | 1,632 tok/s PP512 | 1,109.7 MiB | basecompute | View Benchmark |
Qwen3.8-27B-Q4-mtp BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | Apple M4 Max Metal | 31.5 tok/s TG128 @ 1 ctx | 221 tok/s PP512 | 22,666.1 MiB | basecompute | View Benchmark |
Qwen3-30B-A3B-Instruct-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 139.8 tok/s TG128 @ 1 ctx | 1,789 tok/s PP512 | 1,470.0 MiB | basecompute | View Benchmark |
Qwen3-30B-A3B-Thinking-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 140.1 tok/s TG128 @ 1 ctx | 1,797 tok/s PP512 | 1,462.4 MiB | basecompute | View Benchmark |
gemma-4-26B-A4B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 81.5 tok/s TG128 @ 1 ctx | 1,664 tok/s PP512 | 3,451.3 MiB | basecompute | View Benchmark |
Qwen3.6-27B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 30.3 tok/s TG128 @ 1 ctx | 221 tok/s PP512 | 22,666.8 MiB | basecompute | View Benchmark |
gpt-oss-20b-MXFP4 BaseRTmxfp4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 146.7 tok/s TG128 @ 1 ctx | 1,715 tok/s PP512 | 10,483.6 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 146.9 tok/s TG128 @ 1 ctx | 1,700 tok/s PP512 | 8,231.6 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 160.5 tok/s TG128 @ 1 ctx | 1,709 tok/s PP512 | 7,879.2 MiB | basecompute | View Benchmark |
Qwen3-8B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 59.6 tok/s TG128 @ 1 ctx | 942 tok/s PP512 | 10,364.5 MiB | basecompute | View Benchmark |
gemma-4-E4B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 77.4 tok/s TG128 @ 1 ctx | 2,481 tok/s PP512 | 5,526.5 MiB | basecompute | View Benchmark |
Llama-3.1-8B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 60.8 tok/s TG128 @ 1 ctx | 947 tok/s PP512 | 10,321.6 MiB | basecompute | View Benchmark |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 64.1 tok/s TG128 @ 1 ctx | 946 tok/s PP512 | 10,258.4 MiB | basecompute | View Benchmark |
gemma-4-E2B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 132.2 tok/s TG128 @ 1 ctx | 7,448 tok/s PP512 | 2,611.3 MiB | basecompute | View Benchmark |
Llama-3.1-8B-Instruct-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 105.5 tok/s TG128 @ 1 ctx | 946 tok/s PP512 | 6,176.5 MiB | basecompute | View Benchmark |
Qwen3-8B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 99.2 tok/s TG128 @ 1 ctx | 943 tok/s PP512 | 6,258.2 MiB | basecompute | View Benchmark |
gemma-4-E4B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 109.3 tok/s TG128 @ 1 ctx | 2,552 tok/s PP512 | 3,326.8 MiB | basecompute | View Benchmark |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 110.3 tok/s TG128 @ 1 ctx | 947 tok/s PP512 | 6,114.6 MiB | basecompute | View Benchmark |
Qwen3-4B-Thinking-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 104.3 tok/s TG128 @ 1 ctx | 1,714 tok/s PP512 | 6,036.0 MiB | basecompute | View Benchmark |
Qwen3-4B-Instruct-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 104.3 tok/s TG128 @ 1 ctx | 1,723 tok/s PP512 | 6,035.8 MiB | basecompute | View Benchmark |