Amd64 Family 23 Model 104 Stepping 1 Authenticamd LLM benchmark
The highest recorded decode rate in these Amd64 Family 23 Model 104 Stepping 1 Authenticamd submissions is from llama3.2 (1.2B) at 23.9 decode tok/s via ollama (source run). This page contains 4 workload rows from 1 submitted run across 1 model labels. Model sizes, workloads and reported hardware configurations differ; the highest rate is not a quality ranking or a speed guarantee.
Highest recorded decode rate in these submissions
23.9 decode tok/s
llama3.2 (1.2B) via ollama (Q8_0). Workload: chat-short. Reported hardware: AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB. see full run
llama3.2
| Workload | Model size | Reported hardware | Backend | Quant | decode tok/s | prefill tok/s | TTFT | Run |
|---|---|---|---|---|---|---|---|---|
| chat-short | 1.2B | AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB | ollama@0.34.0 | Q8_0 | 23.88tok/s | 39.41tok/s | 3,172ms | r_b1rn5gi6b3b |
| chat-long | 1.2B | AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB | ollama@0.34.0 | Q8_0 | 14.69tok/s | 93.91tok/s | 33,563ms | r_b1rn5gi6b3b |
| concurrent-decode | 1.2B | AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB | ollama@0.34.0 | Q8_0 | 19.96tok/s | no data | no data | r_b1rn5gi6b3b |
| agent-trace | 1.2B | AMD64 Family 23 Model 104 Stepping 1, AuthenticAMD (6c) + 15GB | ollama@0.34.0 | Q8_0 | 16.80tok/s | 203.2tok/s | 7,977ms | r_b1rn5gi6b3b |
Models measured on Amd64 Family 23 Model 104 Stepping 1 Authenticamd
Common questions about Amd64 Family 23 Model 104 Stepping 1 Authenticamd
Direct Q&A drawn from the runs above: fastest LLM, supported model classes, backend rankings, quantization guidance.
Read the Amd64 Family 23 Model 104 Stepping 1 Authenticamd FAQ →