Community performance summary

Llama 3.2 1B × AMD Radeon Pro 5500M

3 signed observations in one performance comparison. Open any run below to inspect its exact configuration and provenance.

3 signed runs · 0 verified · 1 tester
Median prompt1473.96tok/s · 1388.43 to 1602.61
Median generation103.89tok/s · 85.26 to 112.92
Configurations2Q4_K_M · UD-Q5_K_XL
Last observedSep 7RDNA 1

What the results show

One clear view.
Every recorded run.

Measurements from comparable hardware and models are presented together. Medians show typical performance, while every signed run remains available below for inspection and sharing.

Quantizations observed2 groups
Q4_K_M2 runsUD-Q5_K_XL1 runs
macOS

macOS 26.5.2 Tahoe · macOS 26.6.2 Tahoe

Recorded evidence

Inspect every run.

Open any result to review its exact hardware, model configuration, measurements and submission details.

1 to 8 of 3 configurations

Verified benchmark Community submission
Model / quantHardwareArchitecturePromptGenerationSource
Llama 3.2 1BQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
RDNA 1
1473.96tok/s
112.92tok/s
mnerdSep 7, 2026
Llama 3.2 1BQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
RDNA 1
1602.61tok/s
103.89tok/s
mnerdAug 31, 2026
Llama 3.2 1BUD-Q5_K_XL · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
RDNA 1
1388.43tok/s
85.26tok/s
mnerdJul 31, 2026