← Back to feed

Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things

L4 · DeveloperModels & ReleasesSimon Willison's Blog· 8/16/2026

Developer-focused analysis of a new model's performance characteristics and practical runtime considerations for local deployment.

AI Summary

Qwen 3.8 27B is a new open-source vision-capable LLM that defaults to an 'xhigh' reasoning effort setting, causing extensive overthinking even for simple tasks.

Excerpt

Heees my review of Qwen 3.8 27B - I can't remember the last time I've had this much fun playing with a local model that runs on my own computers simonwillison.net/2026/Aug/16/...

Read Original
0 upvotes · 0 downvotes · 1 min read

Related Articles

L4 · DeveloperModels & Releases@GoogleDeepMind
Google DeepMind (@GoogleDeepMind): We’re rolling out Gemini Omni 1.1 Flash to make generative video highly controllable, faster to iterate on, and more polished for production-grade use. Here’s how you can try it in @FlowbyGoogle and…

Google DeepMind is releasing Gemini Omni 1.1 Flash, a generative video model focused on controllability, faster iteration, and production readiness.

L4 · DeveloperModels & ReleasesHacker News
Gemini-3.5-Transcribe

Google launches Gemini 3.5 Transcribe, a speech-to-text model with 4.0% WER for streaming and 2.6% for non-streaming use cases, now available via API.

L3 · BuilderModels & ReleasesHacker News
Gemini Omni 1.1 Flash

Google released Gemini Omni 1.1 Flash, a production-ready video generation model with new controls for scene extension, frame interpolation, 4K upscaling, and faster prototyping.

L3 · BuilderModels & ReleasesSimon Willison's Blog
Qwen3.8-Flash-Next

Qwen3.8-Flash-Next is a new 125B token multimodal Mixture-of-Experts model that serves as an early preview of the Qwen4 architecture, with only 6B active parameters for performance efficiency.