Skip to content
llm-speed

M3 Max vs M4 Max

Signed, community-submitted decode tok/s, prefill, and TTFT for M3 Max vs M4 Max — every number links to the run it came from.

Verdict

M3 Max has no submitted run yet, so there is no head-to-head number to report. We do have 2 models measured for M4 Max (peak 113 tok/s decode on qwen3-coder-bench-32k) — the table below is that baseline. Run the llm-speed suite on M3 Max and this page fills in the second column automatically.

hardware
M3 Max
Apple M3 · up to 128 GB unified memory
View M3 Max page →
hardware
M4 Max
Apple M4 · up to 128 GB unified memory
View M4 Max page →

M4 Max benchmarks — no shared model with M3 Max yet

Modeldecode tok/sWorkloadRun
qwen3-coder-bench-32k113.3tok/schat-shortr_roktphpc--8
qwen3-coder109.9tok/sconcurrent-decoder_r0di2hkku1h

See also: M3 Max benchmarks · M4 Max benchmarks · All hardware · All models · Methodology