Open-source LLMs
14 articles tagged with this topic
Alibaba and Zhipu Bet on Small Models — Local AI Faces Choice Overload
Qwen Flash and GLM Flash launched together, leaving local users with choice overload. China's open-source LLMs shift from parameter wars to same-tier
Two 3090s Run 27B Model at 165 tok/s — Local AI Is Finally 'Good Enough'
We noted a Reddit user hit 165 tok/s on a 27B Qwen model using two RTX 3090s (used rig under ¥20K) — Agent-ready. Local LLMs just crossed from 'toy' t
GLM, Qwen Catch Closed-Source Leaders on Agent Benchmarks in Just Two Months
GLM and Qwen now match top closed-source models on Agent Arena Code — a result unthinkable two months ago.
BT instead of HuggingFace? Reddit user pitches it — three big hurdles
50GB+ open-source models strain HuggingFace's bandwidth budget. Reddit's BT-sharing idea sounds win-win — but security, dead torrents, and business vi
Developers Turn to BitTorrent Over Nvidia's Hugging Face Takeover
Hugging Face faces Nvidia acquisition rumors; community fears closure. Reddit users note AI models aren't pirated content and can be legally torrented
Zhipu flagship priced at 1/40 of Opus 4.8 — Chinese AI rewrites the default
GLM-5.3-Flash ties Opus 4.8 at 57 on Artificial Analysis, priced at 1/40th. First Chinese model combining frontier performance, low cost, and MIT lice
Hugging Face Up for Sale at $13B — Will Open-Source AI's Core Hub Change Hands?
Hugging Face is reportedly in sale talks at ~$13B. If closed, the deal could reshape open-source AI's ecosystem and corporate AI defaults.
AI's Hottest Post This Week: 8 Words — Hype Is Now Meme-Driven
r/LocalLLaMA's top post: just "It's here!" + a Simpsons GIF, poster ID echoing Musk. Zero tech detail, viral anyway — worth more than any launch.
An 'I have a problem' empty post on LocalLLaMA is itself an industry signal
An empty 'I have a problem' post hit r/LocalLLaMA — a signal-density shift in the open-source LLM community. Non-developers can skip it.
Zhipu GLM's 'Hilarious' Thinking Goes Viral as Chinese Open LLMs Race on Inner Monologue
Zhipu GLM 5.3's 'hilarious' thinking hit r/LocalLLaMA. The meme masks Chinese open LLMs selling transparent reasoning as a post-R1 differentiator.
Qwen 2.4T hits 350 tokens/sec — a price war for large models is coming
Qwen 2.4T hit 350 tokens/sec on NVIDIA's GB300 rack. Open-source is crushing inference costs — a 2026 China cloud price war looks likely.
One Deleted Post on LocalLLaMA Exposes Open-Source AI's Governance Crisis
This week on r/LocalLLaMA, an 11-word 'Why my post was deleted?' post sparked unexpected debate — a signal that open-source AI governance is cracking.
Qwen3.8 27B Runs 200 Tokens/Sec on a Single GPU, Closing Gap with Cloud APIs
Qwen3.8 27B with NInfer hits ~200 tokens/sec on a consumer RTX 5090. Local AI hardware barriers are falling fast, with real implications for enterpris
Three Open LLMs Dropped Same Day — The 'Frontier' Shelf Life Drops Below One Month
r/LocalLLaMA dubbed an ordinary Tuesday "Models Day" — at least three locally-runnable open-source LLMs landed in a single day. The gap between fronti