Chips
Loading community benchmarks…
Loading community benchmarks…
Submitted benchmarks for NVIDIA H100 80GB HBM3 × 8. Browse individual runs and their measurement settings, or open the report JSON for full benchmark details.
Peak throughput is the highest reported value and may come from different runs. Compare model, quantisation, backend, and token counts before drawing conclusions. Model verification identifies published artifact bytes; community benchmark execution is not remotely attested.
NVIDIA H100 80GB HBM3 × 8 CUDA |
95.3 tok/s TG128 |
2,513 tok/s PP512 |
| 17,724.0 MiB |
| sarthak247 |
| View Benchmark |