Models
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Qwen3 8B. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
| Date |
|---|
| Report |
|---|
basecompute/Qwen3-8B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 56.9 tok/s TG128 | 1,707 tok/s PP512 | 4,423.3 MiB | isu | View Benchmark | |
basecompute/Qwen3-8B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 60.6 tok/s TG128 | 1,763 tok/s PP512 | 6,270.6 MiB | isu | View Benchmark | |
basecompute/Qwen3-8B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 60.5 tok/s TG128 | 1,763 tok/s PP512 | 6,270.0 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-8B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 35.2 tok/s TG128 | 1,748 tok/s PP512 | 10,376.1 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ2_K The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 72.8 tok/s TG128 | 1,477 tok/s PP512 | 5,611.1 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ2_K The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 69.6 tok/s TG128 | 1,465 tok/s PP512 | 5,825.1 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ3_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 61.5 tok/s TG128 | 1,418 tok/s PP512 | 6,415.0 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ3_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 60.9 tok/s TG128 | 1,414 tok/s PP512 | 6,588.2 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ4_1 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 51.9 tok/s TG128 | 1,535 tok/s PP512 | 7,485.8 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 55.0 tok/s TG128 | 1,401 tok/s PP512 | 7,276.4 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 53.6 tok/s TG128 | 1,396 tok/s PP512 | 7,379.1 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ5_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 47.9 tok/s TG128 | 1,345 tok/s PP512 | 8,061.1 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ5_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 47.6 tok/s TG128 | 1,353 tok/s PP512 | 8,087.2 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ6_K The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 42.6 tok/s TG128 | 1,409 tok/s PP512 | 8,893.8 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ6_K The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 39.4 tok/s TG128 | 1,419 tok/s PP512 | 9,624.7 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 34.4 tok/s TG128 | 1,473 tok/s PP512 | 10,786.9 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 29.2 tok/s TG128 | 1,448 tok/s PP512 | 12,808.3 MiB | arki05 | View Benchmark | |
Qwen3-8B llama.cppBF16 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 18.9 tok/s TG128 | 1,376 tok/s PP512 | 18,109.9 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-8B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 58.2 tok/s TG128 | 1,758 tok/s PP512 | 6,266.1 MiB | isu | View Benchmark | |
basecompute/Qwen3-8B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 59.7 tok/s TG128 | 1,761 tok/s PP512 | 6,206.0 MiB | isu | View Benchmark | |
Qwen3-8B-base-q2 BaseRTQ2 This report predates artifact identity, so ComputeArena mapped its model name to a model family by hand. The exact model bytes were not verified. | Apple M5 Pro Metal | 81.8 tok/s TG128 | 1,786 tok/s PP512 | 4,615.7 MiB | arki05 | View Benchmark | |
Qwen3-8B-base-q3 BaseRTQ3 This report predates artifact identity, so ComputeArena mapped its model name to a model family by hand. The exact model bytes were not verified. | Apple M5 Pro Metal | 62.0 tok/s TG128 | 1,748 tok/s PP512 | 5,710.4 MiB | arki05 | View Benchmark | |
Qwen3-8B-base-q5 BaseRTQ5 This report predates artifact identity, so ComputeArena mapped its model name to a model family by hand. The exact model bytes were not verified. | Apple M5 Pro Metal | 48.4 tok/s TG128 | 1,730 tok/s PP512 | 7,353.1 MiB | arki05 | View Benchmark | |
Qwen3-8B-base-q6 BaseRTQ6 This report predates artifact identity, so ComputeArena mapped its model name to a model family by hand. The exact model bytes were not verified. | Apple M5 Pro Metal | 42.1 tok/s TG128 | 1,723 tok/s PP512 | 8,448.1 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-8B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 60.7 tok/s TG128 | 1,755 tok/s PP512 | 6,269.8 MiB | arki05 | View Benchmark |