QwenDeepSeek
Qwen Users Roast Reasoning Models: 50% of 'Thinking' Is Just 'Wait' Tokens
Reddit joke exposes a real problem: reasoning models' thinking chains are filled with filler like 'wait', bloating KV cache and exploding deployment c
Aug 18·2 min read