Community performance summary

Models Zai Org GLM 4.6V Flash

1 GPU configurations measured with this model. Compare hardware differences, quantizations and every recorded result.

1 signed run · 0 verified · 1 tester
Median prompt1131.82tok/s · 1131.82 to 1131.82
Median generation72.12tok/s · 72.12 to 72.12
Configurations1Q4_K_M
Last observedAug 6RDNA

What the results show

One clear view.
Every recorded run.

Measurements from comparable hardware and models are presented together. Medians show typical performance, while every signed run remains available below for inspection and sharing.

Hardware tested1 groups
AMD Radeon GFX10, 16 GB1 run
macOS

macOS 15.7.8 Sequoia

Recorded evidence

Inspect every run.

Open any result to review its exact hardware, model configuration, measurements and submission details.

1 to 8 of 1 configurations

Verified benchmark Community submission
Model / quantHardwareArchitecturePromptGenerationSource
Models Zai Org GLM 4.6V FlashQ4_K_M · Dense · Signed ToshLLM app · amd-gpu · Metal · comparison
RDNA
1131.82tok/s
72.12tok/s
ToshLLM communityAug 6, 2026