Skip to content
llm-speed
Leaderboard/hardware/amd64-family-23-model-104-stepping-1-authenticamd

Amd64 Family 23 Model 104 Stepping 1 Authenticamd LLM benchmark

The highest recorded decode rate in these Amd64 Family 23 Model 104 Stepping 1 Authenticamd submissions is from llama3.2 (1.2B) at 23.9 decode tok/s via ollama (source run). This page contains 4 workload rows from 1 submitted run across 1 model labels. Model sizes, workloads and reported hardware configurations differ; the highest rate is not a quality ranking or a speed guarantee.

Highest recorded decode rate in these submissions

23.9 decode tok/s

llama3.2 (1.2B) via ollama (Q8_0). Workload: chat-short. Reported hardware: AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB. see full run

llama3.2

WorkloadModel sizeReported hardwareBackendQuantdecode tok/sprefill tok/sTTFTRun
chat-short1.2B
AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB
ollama@0.34.0Q8_023.88tok/s39.41tok/s3,172msr_b1rn5gi6b3b
chat-long1.2B
AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB
ollama@0.34.0Q8_014.69tok/s93.91tok/s33,563msr_b1rn5gi6b3b
concurrent-decode1.2B
AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB
ollama@0.34.0Q8_019.96tok/sno datano datar_b1rn5gi6b3b
agent-trace1.2B
AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB
ollama@0.34.0Q8_016.80tok/s203.2tok/s7,977msr_b1rn5gi6b3b

Models measured on Amd64 Family 23 Model 104 Stepping 1 Authenticamd

Common questions about Amd64 Family 23 Model 104 Stepping 1 Authenticamd

Direct Q&A drawn from the runs above: fastest LLM, supported model classes, backend rankings, quantization guidance.

Read the Amd64 Family 23 Model 104 Stepping 1 Authenticamd FAQ →