Models
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Gemma 4 E2B IT. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
| Date |
|---|
| Report |
|---|
basecompute/gemma-4-E2B-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 144.6 tok/s TG128 | 13,240 tok/s PP512 | 1,590.9 MiB | arki05 | View Benchmark | |
basecompute/gemma-4-E2B-it BaseRTQ8· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 96.0 tok/s TG128 | 12,646 tok/s PP512 | 2,617.1 MiB | arki05 | View Benchmark | |
basecompute/gemma-4-E2B-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Metal | 48.7 tok/s TG128 | 3,484 tok/s PP512 | 903.1 MiB | skogul97 | View Benchmark | |
basecompute/gemma-4-E2B-it BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Max Metal | 197.5 tok/s TG128 | 20,099 tok/s PP512 | 1,624.1 MiB | lukas | View Benchmark | |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | NVIDIA GB10 CUDA | 139.8 tok/s TG128 @ 1 ctx | 12,036 tok/s PP512 | 3,680.3 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-cuda-q4mix BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | NVIDIA GB10 CUDA | 79.0 tok/s TG128 @ 1 ctx | 18,372 tok/s PP512 | 4,988.4 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-cuda-q8 BaseRTQ4 Nothing in this report ties it to a published model file. Reports from CLI 0.1.0 record only a model name and a file hash, so the run is grouped by the name it reported until it is reconciled. | NVIDIA GB10 CUDA | 87.7 tok/s TG128 @ 1 ctx | 19,246 tok/s PP512 | 5,786.0 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 175.3 tok/s TG128 @ 1 ctx | 7,670 tok/s PP512 | 1,585.0 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 132.2 tok/s TG128 @ 1 ctx | 7,448 tok/s PP512 | 2,611.3 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 83.9 tok/s TG128 @ 1 ctx | 3,628 tok/s PP512 | 2,611.0 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 91.1 tok/s TG128 @ 1 ctx | 4,442 tok/s PP512 | 2,611.6 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 120.8 tok/s TG128 @ 1 ctx | 4,521 tok/s PP512 | 1,587.5 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q8 BaseRTQ8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M3 Ultra Metal | 116.6 tok/s TG128 @ 1 ctx | 8,610 tok/s PP512 | 2,608.4 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M3 Ultra Metal | 137.4 tok/s TG128 @ 1 ctx | 8,996 tok/s PP512 | 1,581.7 MiB | basecompute | View Benchmark | |
gemma-4-E2B-it-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 126.0 tok/s TG128 @ 1 ctx | 3,722 tok/s PP512 | 1,586.5 MiB | basecompute | View Benchmark |