Agents & InferenceSimon Willison

Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

The newly released Qwen 3.8 27B model defaults to an "extra high" reasoning setting that can consume over 22,000 reasoning tokens for simple prompts, easily exceeding standard 8,192-token context limits and spiking generation times to over 20 minutes. If you deploy this model in production pipelines, you must explicitly override the default reasoning effort parameter or expand your context window to avoid catastrophic latency bottlenecks and immediate runner failures on routine tasks.