Models
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for Mistral 7B Instruct v0.3. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
| Date |
|---|
| Report |
|---|
basecompute/Mistral-7B-Instruct-v0.3 BaseRTQ4· default-q4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 66.0 tok/s TG128 | 1,795 tok/s PP512 | 6,126.0 MiB | arki05 | View Benchmark | |
basecompute/Mistral-7B-Instruct-v0.3 BaseRTQ4· default-q8 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M5 Pro Metal | 37.4 tok/s TG128 | 1,783 tok/s PP512 | 10,269.9 MiB | arki05 | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | NVIDIA GB10 CUDA | 53.5 tok/s TG128 @ 1 ctx | 7,554 tok/s PP512 | 4,470.8 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | NVIDIA GB10 CUDA | 27.0 tok/s TG128 @ 1 ctx | 7,510 tok/s PP512 | 7,590.8 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 110.3 tok/s TG128 @ 1 ctx | 947 tok/s PP512 | 6,114.6 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Max Metal | 64.1 tok/s TG128 @ 1 ctx | 946 tok/s PP512 | 10,258.4 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 33.8 tok/s TG128 @ 1 ctx | 393 tok/s PP512 | 10,262.8 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 66.4 tok/s TG128 @ 1 ctx | 511 tok/s PP512 | 6,120.6 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M4 Pro Metal | 59.0 tok/s TG128 @ 1 ctx | 394 tok/s PP512 | 6,119.0 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M1 Max Metal | 40.6 tok/s TG128 @ 1 ctx | 509 tok/s PP512 | 10,262.8 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q8 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M3 Ultra Metal | 78.9 tok/s TG128 @ 1 ctx | 1,388 tok/s PP512 | 10,257.2 MiB | basecompute | View Benchmark | |
Mistral-7B-Instruct-v0.3-Q4 BaseRTQ4 The report's model SHA-256 matches a file published on Hugging Face, so the exact model bytes are known. This does not attest benchmark execution. | Apple M3 Ultra Metal | 129.3 tok/s TG128 @ 1 ctx | 1,440 tok/s PP512 | 6,113.4 MiB | basecompute | View Benchmark |