Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M1 Max. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M1 Max Metal |
17.9 tok/s TG128 @ 1 ctx |
80 tok/s PP512 |
| 22,669.0 MiB |
| basecompute |
| View Benchmark |
gpt-oss-20b-MXFP4 BaseRTmxfp4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 76.6 tok/s TG128 @ 1 ctx | 890 tok/s PP512 | 10,485.2 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 76.7 tok/s TG128 @ 1 ctx | 880 tok/s PP512 | 8,233.5 MiB | basecompute | View Benchmark |
gpt-oss-20b-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 82.3 tok/s TG128 @ 1 ctx | 881 tok/s PP512 | 7,880.5 MiB | basecompute | View Benchmark |
Qwen3-8B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 38.1 tok/s TG128 @ 1 ctx | 500 tok/s PP512 | 10,365.4 MiB | basecompute | View Benchmark |
gemma-4-E4B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 52.0 tok/s TG128 @ 1 ctx | 1,368 tok/s PP512 | 5,526.1 MiB | basecompute | View Benchmark |
Llama-3.1-8B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 38.6 tok/s TG128 @ 1 ctx | 509 tok/s PP512 | 10,323.0 MiB | basecompute | View Benchmark |
Qwen3-8B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 60.2 tok/s TG128 @ 1 ctx | 507 tok/s PP512 | 6,259.0 MiB | basecompute | View Benchmark |
gemma-4-E4B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 75.9 tok/s TG128 @ 1 ctx | 1,394 tok/s PP512 | 3,326.7 MiB | basecompute | View Benchmark |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 66.4 tok/s TG128 @ 1 ctx | 511 tok/s PP512 | 6,120.6 MiB | basecompute | View Benchmark |
Qwen3-4B-Instruct-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 72.1 tok/s TG128 @ 1 ctx | 925 tok/s PP512 | 6,038.4 MiB | basecompute | View Benchmark |
Llama-3.2-3B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 88.4 tok/s TG128 @ 1 ctx | 1,240 tok/s PP512 | 4,774.1 MiB | basecompute | View Benchmark |
Qwen3-4B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 71.3 tok/s TG128 @ 1 ctx | 920 tok/s PP512 | 6,036.9 MiB | basecompute | View Benchmark |
Qwen3-4B-Thinking-2507-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 71.5 tok/s TG128 @ 1 ctx | 925 tok/s PP512 | 6,036.8 MiB | basecompute | View Benchmark |
Llama-3.1-8B-Instruct-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 63.7 tok/s TG128 @ 1 ctx | 509 tok/s PP512 | 6,179.6 MiB | basecompute | View Benchmark |
gemma-4-E2B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 91.1 tok/s TG128 @ 1 ctx | 4,442 tok/s PP512 | 2,611.6 MiB | basecompute | View Benchmark |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 40.6 tok/s TG128 @ 1 ctx | 509 tok/s PP512 | 10,262.8 MiB | basecompute | View Benchmark |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 120.8 tok/s TG128 @ 1 ctx | 4,521 tok/s PP512 | 1,587.5 MiB | basecompute | View Benchmark |
Qwen3-4B-Thinking-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 93.2 tok/s TG128 @ 1 ctx | 921 tok/s PP512 | 3,900.1 MiB | basecompute | View Benchmark |
Qwen3-4B-Instruct-2507-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 91.3 tok/s TG128 @ 1 ctx | 922 tok/s PP512 | 3,898.2 MiB | basecompute | View Benchmark |
Llama-3.2-3B-Instruct-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 110.7 tok/s TG128 @ 1 ctx | 1,229 tok/s PP512 | 3,091.7 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Base-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 140.7 tok/s TG128 @ 1 ctx | 438 tok/s PP512 | 2,955.2 MiB | basecompute | View Benchmark |
Qwen3.5-2B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 141.5 tok/s TG128 @ 1 ctx | 400 tok/s PP512 | 2,953.7 MiB | basecompute | View Benchmark |
Qwen3-4B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 101.9 tok/s TG128 @ 1 ctx | 927 tok/s PP512 | 3,899.5 MiB | basecompute | View Benchmark |
Llama-3.2-1B-Instruct-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 229.4 tok/s TG128 @ 1 ctx | 3,372 tok/s PP512 | 1,676.2 MiB | basecompute | View Benchmark |