Models
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Qwen3 0.6B. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
| Date |
|---|
| Report |
|---|
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Pro Metal | 268.8 tok/s TG128 @ 1 ctx | 2,847 tok/s PP512 | 1,362.6 MiB | isu | View Benchmark | |
Qwen3-0.6B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M3 Ultra Metal | 440.2 tok/s TG128 @ 1 ctx | 12,494 tok/s PP512 | 1,641.7 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M3 Ultra Metal | 501.7 tok/s TG128 @ 1 ctx | 12,830 tok/s PP512 | 1,368.9 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 284.7 tok/s TG128 @ 1 ctx | 4,787 tok/s PP512 | 1,645.8 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 455.8 tok/s TG128 @ 1 ctx | 9,749 tok/s PP512 | 1,643.9 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 208.6 tok/s TG128 @ 1 ctx | 3,571 tok/s PP512 | 1,366.9 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 74.3 tok/s TG128 @ 1 ctx | 2,225 tok/s PP512 | 1,645.7 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | NVIDIA GB10 CUDA | 458.8 tok/s TG128 @ 1 ctx | 20,970 tok/s PP512 | 1,041.2 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-cuda-q4 BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | NVIDIA GB10 CUDA | 456.1 tok/s TG128 @ 1 ctx | 52,750 tok/s PP512 | 1,236.9 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-cuda-q4mix BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | NVIDIA GB10 CUDA | 413.2 tok/s TG128 @ 1 ctx | 49,865 tok/s PP512 | 1,130.4 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 583.3 tok/s TG128 @ 1 ctx | 9,998 tok/s PP512 | 1,371.1 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 100.8 tok/s TG128 @ 1 ctx | 1,851 tok/s PP512 | 1,373.0 MiB | basecompute | View Benchmark | |
Qwen3-0.6B-cuda-q8 BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | NVIDIA GB10 CUDA | 315.8 tok/s TG128 @ 1 ctx | 51,373 tok/s PP512 | 1,359.7 MiB | basecompute | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 481.5 tok/s TG128 @ 1 ctx | 19,595 tok/s PP512 | 1,367.1 MiB | isu | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 698.5 tok/s TG128 | 32,977 tok/s PP512 | 1,393.8 MiB | lukas | View Benchmark | |
Qwen3-0.6B llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | AMD Radeon RX 7900 XT (RADV NAVI31) Vulkan | 479.4 tok/s TG128 | 26,346 tok/s PP512 | 291.1 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | AMD Radeon RX 7900 XT (RADV NAVI31) Vulkan | 491.9 tok/s TG128 | 25,964 tok/s PP512 | 258.8 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | AMD Radeon RX 7900 XT (RADV NAVI31) Vulkan | 508.8 tok/s TG128 | 26,137 tok/s PP512 | 270.9 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | AMD Radeon RX 7900 XT ROCm | 313.7 tok/s TG128 | 24,691 tok/s PP512 | 2,455.8 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | AMD Radeon RX 7900 XT ROCm | 367.7 tok/s TG128 | 23,718 tok/s PP512 | 2,222.3 MiB | arki05 | View Benchmark | |
Qwen3-0.6B llama.cppQ4_K_M The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | AMD Radeon RX 7900 XT ROCm | 368.6 tok/s TG128 | 23,972 tok/s PP512 | 2,358.9 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 350.2 tok/s TG128 | 20,537 tok/s PP512 | 1,653.5 MiB | arki05 | View Benchmark | |
basecompute/Qwen3-0.6B BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 516.8 tok/s TG128 | 20,778 tok/s PP512 | 1,380.5 MiB | arki05 | View Benchmark | |
Qwen/Qwen3-0.6B BaseRTQ4· default-q4 This report predates artifact identity, so ComputeArena mapped its model name to a model family by hand. The exact model bytes were not verified. | Apple M5 Metal | 172.7 tok/s TG128 | 4,997 tok/s PP512 | 1,389.9 MiB | skogul97 | View Benchmark | |
Qwen3-0.6B llama.cppQ8_0 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro BLAS + Metal | 283.0 tok/s TG128 | 14,942 tok/s PP512 | 2,565.6 MiB | arki05 | View Benchmark |