TECH FLOW Svět Androida
← Back to the stream
simonwillison.net · picked by Petr Mišák · 47d ago

Qwen 3.8 27B: powerful model, but default mode leads to wild overthinking

Source preview: Qwen 3.8 27B: powerful model, but default mode leads to wild overthinking
AI summary

Qwen 3.8 27B is a powerful open-source model from Alibaba suitable for local deployment, but its default setting automatically engages in excessively deep analysis (xhigh reasoning mode). The author recommends running the model at lower reasoning levels to avoid wasting time on unnecessary computation.

The summary is written by AI from the source; it isn’t the newsroom’s opinion. For details, read the source.

8 people have already opened the source

Tip author’s note

It's nice to see how quantization and model size fundamentally affect output quality.

AI questions & answers
What are typical deployments for the Qwen 3.8 27B model?

Given its 27 billion parameters and 17GB GGUF version, the model runs well on adequately-resourced laptops or small servers. It is commonly deployed via LM Studio or llama-server.

What is the xhigh reasoning mode in Qwen models and why is it problematic?

The xhigh (extra high) mode increases the depth of the model's reasoning for complex tasks. While useful for analytical problems, it leads to significant time waste for simple requests like drawing a circle without corresponding benefit.

What vision capabilities does the Qwen 3.8 27B model demonstrate?

The model performs well on detection tasks, such as returning precise bounding boxes for objects in photographs in a requested scale. It shows competency in multimodal inputs including images.

Questions and answers are written by AI about the topic, not taken from the source; they aren’t the newsroom’s opinion.
Related from the stream
Mentions