Qwen 3.8 27B defaults to excessive reasoning effort causing slow local inference
Why it matters — Engineers running the model locally must override the default to avoid multi-minute waits for trivial prompts. The setting also risks exhausting context windows on modest hardware, limiting practical use cases without manual tuning.
↗