Community performance summary

Openai GPT OSS 120B

1 GPU configurations measured with this model. Compare hardware differences, quantizations and every recorded result.

1 signed run · 0 verified · 1 tester
Median prompt266.48tok/s · 266.48 to 266.48
Median generation65.99tok/s · 65.99 to 65.99
Configurations1MXFP4
Last observedSep 14GCN / Vega + RDNA 2

What the results show

One clear view.
Every recorded run.

Measurements from comparable hardware and models are presented together. Medians show typical performance, while every signed run remains available below for inspection and sharing.

Hardware tested1 groups
AMD Radeon Pro Vega II + AMD Radeon PRO W6800X Duo1 run
macOS

macOS 26.6.2 Tahoe

Recorded evidence

Inspect every run.

Open any result to review its exact hardware, model configuration, measurements and submission details.

1 to 8 of 1 configurations

Verified benchmark Community submission
Model / quantHardwareArchitecturePromptGenerationSource
Openai GPT OSS 120BMXFP4 · MoE · Signed ToshLLM app · amd-gpu · Metal · comparison
GCN / Vega + RDNA 2
266.48tok/s
65.99tok/s
threadedSep 14, 2026