Guide
Qwen3.8-27B has the best coding ceiling you can run at home, and it ships on maximum effort
Qwen3.8's chat template sets reasoning effort to xhigh unless you change it. On one Mac that is the difference between an answer in eighty seconds and one in fourteen minutes. We ran all three levels to find out what the slow one buys.
- Qwen3.8-27B posted the highest single coding score in our data from anything that fits on consumer hardware, a 90.7, and at IQ3/Q4 it needs just over 13 GB.
- The default is the ceiling: the chat template resolves reasoning_effort to xhigh when nothing is set, so an untouched install runs every answer at maximum effort.
- Low and medium are the same setting in practice, 4,984 tokens against 4,792 and eighty-four seconds against seventy-seven.
- xhigh costs eight times the tokens and eleven times the wall clock for half a point of median coding score, which is well inside normal run-to-run variation.
- A 12,000-token thinking budget halves the wait without truncating anything, and is the one lever we have tested directly.