Open Source LLMs
6 articles tagged with this topic
Zhipu Open-Sources GLM-5.3-Flash: Nears Claude Opus at One-Tenth the Price
Zhipu open-sources GLM-5.3-Flash: 320B-param MoE with 18B active, claims near-Claude Opus 4.8 performance at one-tenth prior pricing under MIT license
Chinese Open Models Top Global Charts — 'Sleep on the Couch' Meme Is Real
Reddit LocalLLaMA meme about "someone sleeping on the couch tonight" hides real truth: Chinese open models now top LMSys Arena, HuggingFace, and GitHu
Qwen3 Hits 6250 token/s on RTX 5090: Open Source Drops Inference Costs Another 50%
Unsloth's compressed Qwen3 8B hits 6250 token/s on RTX 5090 — 50% faster than traditional Q4, powered by Nvidia's NVFP4 4-bit format.
Local LLaMA Players Brawl Over a New Model — Just Another Day in Open Source
The r/LocalLLaMA community is split again over a new model release. Open-source models' real-world usage often doesn't match their benchmark scores. S
RTX 5080 Runs 30B Model in 2 Minutes—Local AI Catches Up to the Cloud
A consumer GPU runs a 30B model 15x faster than the cloud at near-cloud quality. Local AI is becoming a real enterprise option for data-sensitive indu
IBM Open-Sources Granite 4.1: 21 Quantized Versions Prove Bottleneck Isn't Size
IBM open-sources Granite 4.1. A 21-version quantization test shows no quality difference: small models' bottleneck is base capability, not compression