Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M1 Max. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M1 Max Metal |
52.0 tok/s TG128 @ 1 ctx |
1,368 tok/s PP512 |
| 5,526.1 MiB |
| basecompute |
| View Benchmark |
Qwen3-8B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 38.1 tok/s TG128 @ 1 ctx | 500 tok/s PP512 | 10,365.4 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 82.3 tok/s TG128 @ 1 ctx | 881 tok/s PP512 | 7,880.5 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 76.7 tok/s TG128 @ 1 ctx | 880 tok/s PP512 | 8,233.5 MiB | basecompute | View Benchmark |
gpt-oss-20b-MXFP4 BaseRTmxfp4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 76.6 tok/s TG128 @ 1 ctx | 890 tok/s PP512 | 10,485.2 MiB | basecompute | View Benchmark |
Qwen3.6-27B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 17.9 tok/s TG128 @ 1 ctx | 80 tok/s PP512 | 22,669.0 MiB | basecompute | View Benchmark |
gemma-4-26B-A4B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 51.5 tok/s TG128 @ 1 ctx | 871 tok/s PP512 | 3,451.0 MiB | basecompute | View Benchmark |
Qwen3-30B-A3B-Thinking-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 87.2 tok/s TG128 @ 1 ctx | 901 tok/s PP512 | 1,462.6 MiB | basecompute | View Benchmark |
Qwen3-30B-A3B-Instruct-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 87.7 tok/s TG128 @ 1 ctx | 901 tok/s PP512 | 1,471.3 MiB | basecompute | View Benchmark |
Qwen3.8-27B-Q4-mtp BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | Apple M1 Max Metal | 18.2 tok/s TG128 @ 1 ctx | 82 tok/s PP512 | 22,668.1 MiB | basecompute | View Benchmark |
NVIDIA-Nemotron-3-Nano-30B-A3B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 110.3 tok/s TG128 @ 1 ctx | 870 tok/s PP512 | 1,112.5 MiB | basecompute | View Benchmark |
muse-glimmer-30B-kquant-17gb BaseRTpassthrough_gguf The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 17.3 tok/s TG128 @ 1 ctx | 140 tok/s PP512 | 608.3 MiB | basecompute | View Benchmark |
Qwen3.5-35B-A3B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 93.7 tok/s TG128 @ 1 ctx | 280 tok/s PP512 | 2,750.7 MiB | basecompute | View Benchmark |
Qwen3.6-35B-A3B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 94.2 tok/s TG128 @ 1 ctx | 278 tok/s PP512 | 2,749.9 MiB | basecompute | View Benchmark |
muse-glimmer-30B-kquant-dynamic BaseRTpassthrough_gguf The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 14.3 tok/s TG128 @ 1 ctx | 136 tok/s PP512 | 609.8 MiB | basecompute | View Benchmark |
gemma-4-26B-A4B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 45.5 tok/s TG128 @ 1 ctx | 877 tok/s PP512 | 3,955.2 MiB | basecompute | View Benchmark |
Qwen3.6-27B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 11.0 tok/s TG128 @ 1 ctx | 83 tok/s PP512 | 37,635.9 MiB | basecompute | View Benchmark |
Qwen3.8-27B-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 11.0 tok/s TG128 @ 1 ctx | 82 tok/s PP512 | 37,636.5 MiB | basecompute | View Benchmark |
Qwen3.8-27B-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 11.0 tok/s TG128 @ 1 ctx | 82 tok/s PP512 | 37,636.3 MiB | basecompute | View Benchmark |
Qwen3-30B-A3B-Thinking-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 64.6 tok/s TG128 @ 1 ctx | 893 tok/s PP512 | 1,919.3 MiB | basecompute | View Benchmark |
Qwen3-30B-A3B-Thinking-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 63.7 tok/s TG128 @ 1 ctx | 895 tok/s PP512 | 1,919.4 MiB | basecompute | View Benchmark |
NVIDIA-Nemotron-3-Nano-30B-A3B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 61.1 tok/s TG128 @ 1 ctx | 857 tok/s PP512 | 1,110.5 MiB | basecompute | View Benchmark |
NVIDIA-Nemotron-3-Nano-30B-A3B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 60.6 tok/s TG128 @ 1 ctx | 858 tok/s PP512 | 1,110.8 MiB | basecompute | View Benchmark |
Qwen3.5-35B-A3B-Q8 BaseRTQ8 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | Apple M1 Max Metal | 72.0 tok/s TG128 @ 1 ctx | 250 tok/s PP512 | 59,626.8 MiB | basecompute | View Benchmark |
Qwen3.5-35B-A3B-Q8 BaseRTQ8 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | Apple M1 Max Metal | 72.2 tok/s TG128 @ 1 ctx | 251 tok/s PP512 | 59,578.6 MiB | basecompute | View Benchmark |