Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things
Source
simonwillison.net
Published
Terms in this piece · Glossary
quantization — Shrinking a model by storing its numbers less precisely — like rounding — so it runs faster and fits on smaller hardware, at a small quality cost.
Why it matters
A capable vision model in the 27B range runs on a well-specced laptop, but its default reasoning effort makes it impractical there; anyone trying it should turn that setting down before judging the model.