1 article tagged with this topic
A Reddit dev shrank Qwen 27B to ~23B by stripping layers — no retraining. Open-source Chinese LLMs are now "hackable" for local deployment.