Agents & InferenceSimon Willison

Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Qwen 3.8 27B defaults to "xhigh" reasoning effort, burning 22k+ tokens and 21 minutes to generate a single SVG—20x slower than with reasoning off. This matters because it silently tanks throughput and spikes costs on consumer hardware; if you ship this in production, you’ll need to explicitly cap reasoning effort or risk unpredictable latency and token waste.

LinkedIn

Two AI summaries of each story, blind-voted — see today's agents & inference digest →