Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Apple M5 Max. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
Apple M5 Max Metal |
34.0 tok/s TG128 |
589 tok/s PP512 |
| 14,842.1 MiB |
| basecompute |
| View Benchmark |
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 708.3 tok/s TG128 | 34,136 tok/s PP512 | 1,554.9 MiB | basecompute | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 177.6 tok/s TG128 | 4,527 tok/s PP512 | 1,967.5 MiB | basecompute | View Benchmark |
basecompute/Qwen3-0.6B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 549.5 tok/s TG128 | 33,088 tok/s PP512 | 1,696.0 MiB | basecompute | View Benchmark |
Qwen3 0.6B Instruct llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max BLAS + Metal | 379.0 tok/s TG128 | 24,983 tok/s PP512 | 2,544.5 MiB | basecompute | View Benchmark |
Models Qwen Qwen3 0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max BLAS + Metal | 439.1 tok/s TG128 | 24,365 tok/s PP512 | 2,396.2 MiB | basecompute | View Benchmark |
basecompute/Qwen3.8-27B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 19.7 tok/s TG128 | 589 tok/s PP512 | 22,598.1 MiB | lukas | View Benchmark |
basecompute/Qwen3.6-35B-A3B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 119.6 tok/s TG128 | 1,540 tok/s PP512 | 3,727.4 MiB | lukas | View Benchmark |
basecompute/gemma-4-26B-A4B-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 92.0 tok/s TG128 | 4,055 tok/s PP512 | 3,526.5 MiB | lukas | View Benchmark |
Gemma-4 12B IT (smart Q4_0, QAT-lossless) llama.cppQ4_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max BLAS + Metal | 54.2 tok/s TG128 | 1,777 tok/s PP512 | 7,377.3 MiB | lukas | View Benchmark |
basecompute/Muse-Glimmer-30B BaseRTpassthrough_gguf· default-q4k-dynamic The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 15.7 tok/s TG128 | 830 tok/s PP512 | 657.8 MiB | lukas | View Benchmark |
basecompute/Qwen3.6-27B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 31.3 tok/s TG128 | 603 tok/s PP512 | 22,605.2 MiB | basecompute | View Benchmark |
basecompute/gpt-oss-20b BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 169.0 tok/s TG128 | 2,179 tok/s PP512 | 7,882.2 MiB | lukas | View Benchmark |
basecompute/gpt-oss-120b BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 114.3 tok/s TG128 | 1,401 tok/s PP512 | 8,325.4 MiB | lukas | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 184.4 tok/s TG128 | 4,984 tok/s PP512 | 1,111.6 MiB | lukas | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 109.2 tok/s TG128 | 4,943 tok/s PP512 | 1,140.9 MiB | lukas | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 183.7 tok/s TG128 | 4,941 tok/s PP512 | 1,140.9 MiB | lukas | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 180.8 tok/s TG128 | 4,951 tok/s PP512 | 1,141.1 MiB | lukas | View Benchmark |
basecompute/NVIDIA-Nemotron-3-Nano-30B-A3B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 177.0 tok/s TG128 | 4,821 tok/s PP512 | 1,141.2 MiB | lukas | View Benchmark |
basecompute/Qwen3.8-27B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 21.5 tok/s TG128 | 590 tok/s PP512 | 22,598.4 MiB | lukas | View Benchmark |
basecompute/Qwen3.8-27B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 20.7 tok/s TG128 | 599 tok/s PP512 | 22,597.8 MiB | lukas | View Benchmark |
basecompute/Qwen3.8-27B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 21.5 tok/s TG128 | 594 tok/s PP512 | 22,598.0 MiB | lukas | View Benchmark |
basecompute/gemma-4-E2B-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 197.5 tok/s TG128 | 20,099 tok/s PP512 | 1,624.1 MiB | lukas | View Benchmark |
basecompute/gemma-4-E4B-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 121.1 tok/s TG128 | 8,156 tok/s PP512 | 3,383.0 MiB | lukas | View Benchmark |
basecompute/gemma-3-1b-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 415.5 tok/s TG128 | 19,770 tok/s PP512 | 954.6 MiB | lukas | View Benchmark |