Community performance summary

Qwen3 Coder 30B-A3B

3 GPU configurations measured with this model. Compare hardware differences, quantizations and every recorded result.

4 signed runs · 0 verified · 3 testers
Median prompt507.17tok/s · 346.78 to 1293.17
Median generation50.56tok/s · 23.94 to 94.98
Configurations1Q4_K_M
Last observedJul 23GCN / Vega · GCN / Vega + RDNA 2 · RDNA 2

What the results show

One clear view.
Every recorded run.

Measurements from comparable hardware and models are presented together. Medians show typical performance, while every signed run remains available below for inspection and sharing.

Hardware tested3 groups
AMD Radeon Vega ×21 runAMD Radeon Pro Vega II + AMD Radeon PRO W6800X2 runsRX 6950 XT1 run
macOS

macOS 15.7.5 Sequoia · macOS 26.5.2 Tahoe

Recorded evidence

Inspect every run.

Open any result to review its exact hardware, model configuration, measurements and submission details.

1 to 8 of 4 configurations

Verified benchmark Community submission
Model / quantHardwareArchitecturePromptGenerationSource
Qwen3 Coder 30B-A3BQ4_K_M · MoE · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
1293.17tok/s
94.98tok/s
gervaiscoquilJul 22, 2026
Qwen3 Coder 30B-A3BQ4_K_M · MoE · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
346.78tok/s
53.32tok/s
gervaiscoquilJul 18, 2026
Qwen3-Coder 30B A3BQ4_K_M · MoE · ncmoe 23 · comparison
RX 6950 XT16 GB VRAM
RDNA 2
587.3tok/s
47.8tok/s
Community testerJul 16, 2026
Qwen3 Coder 30B-A3BQ4_K_M · MoE · Signed ToshLLM app · amd-gpu · Metal · comparison
AMD Radeon Vega ×231.97 GB VRAM
GCN / Vega
427.04tok/s
23.94tok/s
ToshLLM communityJul 23, 2026