Benchmark result
unsloth/Qwen3.8-27B-GGUF on AMD Radeon RX 7900 XTX — 50.0 tok/s
Measured with Unsloth Studio on September 22, 2026.
What the model built
Open full screen →The coding scenario asks for a playable game in a single HTML file. This is exactly what the model returned, unedited; the frame only adds a line that tells this page how tall it is.
How this run compares
11th fastest of 29 runs of this model on AMD Radeon RX 7900 XTX · median 45.8 tok/s.
sort
- 157.4tok/s78.3Unsloth Studio · UD-Q4_K_M · KV q4_0
- 254.9tok/s85.9Unsloth Studio · UD-Q4_K_M · KV q4_0
- 352.1tok/s86.9Unsloth Studio · UD-Q4_K_M · KV q8_0
- 451.7tok/s84.7Unsloth Studio · Q4_K_M · KV q8_0
- 551.7tok/s88.7Unsloth Studio · UD-Q4_K_M · KV q8_0
- 651.2tok/s85.0Unsloth Studio · Q4_K_M · KV f16
- 750.5tok/s89.3Unsloth Studio · UD-Q4_K_M · KV q4_0
- 850.5tok/s89.5Unsloth Studio · UD-Q4_K_M · KV q8_0
- 950.3tok/s87.6Unsloth Studio · UD-Q4_K_M · KV q4_0
- 1050.2tok/s88.1Unsloth Studio · UD-Q4_K_M · KV q4_0
- 1150.0tok/s86.4Unsloth Studio · UD-Q4_K_M · KV q5_0this run
- 1249.4tok/s88.1Unsloth Studio · UD-Q4_K_M · KV q5_0
- 1347.4tok/s89.6Unsloth Studio · Q4_K_M · KV q4_0
- 1447.4tok/s87.9Unsloth Studio · Q4_K_M · KV q5_0
- 1545.8tok/s87.2Unsloth Studio · UD-Q4_K_M · KV q5_0
- 1645.5tok/s84.8Unsloth Studio · Q4_K_M · KV f16
- 1744.5tok/s87.5Unsloth Studio · Q4_K_M · KV q8_0
- 1843.2tok/s80.7Unsloth Studio · Q4_K_M · KV q5_0
- 1942.4tok/s84.9Unsloth Studio · UD-Q4_K_M · KV f16
- 2042.3tok/s87.7Unsloth Studio · UD-Q4_K_M · KV q5_0
- 2140.4tok/s86.1Unsloth Studio · UD-Q4_K_M · KV f16
- 2240.3tok/s87.0Unsloth Studio · UD-Q4_K_M · KV f16
- 2338.8tok/s89.3Unsloth Studio · UD-Q4_K_M · KV f16
- 2438.3tok/s87.1Unsloth Studio · Q4_K_M · KV f16
- 2537.0tok/s91.4Unsloth Studio · Q4_K_M · KV q4_0
- 2635.5tok/s86.4Unsloth Studio · UD-Q4_K_M · KV f16
- 2732.7tok/s90.0Unsloth Studio · Q4_K_M · KV f16
- 2831.3tok/s89.6Unsloth Studio · UD-Q4_K_M · KV q4_0
- 2927.0tok/s88.1Unsloth Studio · KV f16
unsloth/Qwen3.8-27B-GGUF on other hardware
Reproduce this run
toolUnsloth Studio
modelunsloth/Qwen3.8-27B-GGUF (UD-Q4_K_M)
context65,536
kv cacheq5_0
thinkingon
effortxhigh (model default)
think limitnone
samplingtemperature 0.7 · top_p 0.8 · top_k 20 · min_p 0 · presence_penalty 1.5
clientv0.4.80+98
Set those in Unsloth Studio, then:
llm-benchmark benchmark --model "unsloth/Qwen3.8-27B-GGUF" --tool "Unsloth Studio"The client prompts for context length and thinking mode, and for the KV cache dtype on Unsloth Studio. Sampling is left at the model default — the values above are what the tool reported using, not overrides the client sent.