Community performance summary

Llama 3.2 1B × AMD Radeon Pro 555X

2 signed observations in one performance comparison. Open any run below to inspect its exact configuration and provenance.

2 signed runs · 0 verified · 2 testers
Median prompt218.96tok/s · 218.84 to 219.08
Median generation40.14tok/s · 40.06 to 40.21
Configurations1Q8_0
Last observedSep 23GCN / Vega

What the results show

One clear view.
Every recorded run.

Measurements from comparable hardware and models are presented together. Medians show typical performance, while every signed run remains available below for inspection and sharing.

Quantizations observed1 groups
Q8_02 runs
macOS

macOS 15.7.7 Sequoia · macOS 15.7.9 Sequoia

Recorded evidence

Inspect every run.

Open any result to review its exact hardware, model configuration, measurements and submission details.

1 to 8 of 2 configurations

Verified benchmark Community submission
Model / quantHardwareArchitecturePromptGenerationSource
Llama 3.2 1BQ8_0 · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega
218.84tok/s
40.21tok/s
ToshLLM communitySep 23, 2026
Llama 3.2 1BQ8_0 · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega
219.08tok/s
40.06tok/s
toshllm@wbig.gsSep 22, 2026