Qwen 3.8 27B is excellent, but it defaults to overthinking things
Qwen 3.8 27B delivers strong vision, coding, tool use, and long-context performance in a roughly 17GB local model, but its xhigh reasoning default can spend enormous amounts of time on trivial tasks. HN users compare reasoning modes, share llama.cpp and MTP optimizations, and report promising but often slow local coding-agent use.