Back to home
reasoning models
3 articles tagged with this topic
QwenAlibaba
Qwen Makes Thinking Depth Adjustable — Alibaba Lets LLMs Allocate Compute On Demand
Alibaba's Qwen now lets users adjust 'thinking depth'—quick answers for easy questions, more reasoning for hard ones. LLMs shift from on/off switch to
15h ago2 min read
QwenAlibaba
Qwen Thinks 90 Min, No Answer — One Parameter Cures Open-Source Overthinking
Alibaba's Qwen3.8-27b overthinks by default — inferences can run 90 minutes without output. A Reddit dev shared two llamacpp params that fix it.
Aug 172 min read
QwenAlibaba
Qwen3.8 Local Test: One Reasoning Dial, 20x Token Cost — The Hidden Bill
Qwen3.8-27B local test: reasoning_effort swings tokens 20x (2K–40K). "Thinking cost" is the underestimated deployment variable in the reasoning era.
Aug 142 min read