Models
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Qwen3 0.6B. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
| Date |
|---|
| Report |
|---|
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 358.1 tok/s TG128 | 14,344 tok/s PP512 | 2,342.6 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 358.3 tok/s TG128 | 14,509 tok/s PP512 | 2,333.1 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 356.2 tok/s TG128 | 14,552 tok/s PP512 | 2,333.7 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 311.0 tok/s TG128 | 18,474 tok/s PP512 | 1,653.1 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 512.0 tok/s TG128 | 20,603 tok/s PP512 | 1,380.7 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 349.8 tok/s TG128 | 20,366 tok/s PP512 | 1,653.3 MiB | arki05 | View Benchmark | |
Models Qwen Qwen3 0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max BLAS + Metal | 439.1 tok/s TG128 | 24,365 tok/s PP512 | 2,396.2 MiB | basecompute | View Benchmark | |
Qwen3 0.6B Instruct llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max BLAS + Metal | 379.0 tok/s TG128 | 24,983 tok/s PP512 | 2,544.5 MiB | basecompute | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 549.5 tok/s TG128 | 33,088 tok/s PP512 | 1,696.0 MiB | basecompute | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 708.3 tok/s TG128 | 34,136 tok/s PP512 | 1,554.9 MiB | basecompute | View Benchmark |