Community performance summary

Qwen3 4B × Vega ×4 + W6800X Duo ×3

4 signed observations in one performance comparison. Open any run below to inspect its exact configuration and provenance.

4 signed runs · 0 verified · 2 testers
Median prompt542.73tok/s · 532.51 to 794.46
Median generation32.28tok/s · 23.48 to 48.8
Configurations1Q4_K_M
Last observedJul 20GCN / Vega + RDNA 2

What the results show

One clear view.
Every recorded run.

Measurements from comparable hardware and models are presented together. Medians show typical performance, while every signed run remains available below for inspection and sharing.

Quantizations observed1 groups
Q4_K_M4 runs
macOS

macOS 26.5 Tahoe

Recorded evidence

Inspect every run.

Open any result to review its exact hardware, model configuration, measurements and submission details.

1 to 8 of 4 configurations

Verified benchmark Community submission
Model / quantHardwareArchitecturePromptGenerationSource
Qwen3 4BQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
552.59tok/s
48.8tok/s
ToshLLM communityJul 20, 2026
Qwen3 4BQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
794.46tok/s
41.07tok/s
Flint IronstagJul 20, 2026
Qwen3 4BQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
532.51tok/s
23.49tok/s
ToshLLM communityJul 20, 2026
Qwen3 4BQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
532.87tok/s
23.48tok/s
Flint IronstagJul 20, 2026