Skip to content
llm-speed

H100 SXM LLM benchmark

No H100 SXM LLM benchmarks yet. Be the first to submit a signed, reproducible tok/s run from your own H100 SXM.

Will it run local LLMs?

No signed H100 SXM run yet, but its 80GB of VRAM tells you what fits. At 4-bit it comfortably runs models up to about 82B parameters with room for context, and up to roughly 130B with tighter quantization. Check a specific model with the VRAM-fit tool.

No H100 SXM benchmarks yet.

Run on YOUR hardware to populate this page: pipx install llm-speed && llm-speed bench

$ pipx install llm-speed && llm-speed bench

Common questions about H100 SXM

Direct Q&A drawn from the runs above: fastest LLM, supported model classes, backend rankings, quantization guidance.

Read the H100 SXM FAQ →