Back to home

Open Source LLMs

6 articles tagged with this topic

ZhipuZ.ai

Zhipu Open-Sources GLM-5.3-Flash: Nears Claude Opus at One-Tenth the Price

Zhipu open-sources GLM-5.3-Flash: 320B-param MoE with 18B active, claims near-Claude Opus 4.8 performance at one-tenth prior pricing under MIT license

3d ago2 min read
DeepSeekQwen

Chinese Open Models Top Global Charts — 'Sleep on the Couch' Meme Is Real

Reddit LocalLLaMA meme about "someone sleeping on the couch tonight" hides real truth: Chinese open models now top LMSys Arena, HuggingFace, and GitHu

Aug 212 min read
Qwen3Nvidia

Qwen3 Hits 6250 token/s on RTX 5090: Open Source Drops Inference Costs Another 50%

Unsloth's compressed Qwen3 8B hits 6250 token/s on RTX 5090 — 50% faster than traditional Q4, powered by Nvidia's NVFP4 4-bit format.

Aug 212 min read
LocalLLaMAOpen Source LLMs

Local LLaMA Players Brawl Over a New Model — Just Another Day in Open Source

The r/LocalLLaMA community is split again over a new model release. Open-source models' real-world usage often doesn't match their benchmark scores. S

Aug 152 min read
QwenRTX 5080

RTX 5080 Runs 30B Model in 2 Minutes—Local AI Catches Up to the Cloud

A consumer GPU runs a 30B model 15x faster than the cloud at near-cloud quality. Local AI is becoming a real enterprise option for data-sensitive indu

Aug 142 min read
IBMGranite

IBM Open-Sources Granite 4.1: 21 Quantized Versions Prove Bottleneck Isn't Size

IBM open-sources Granite 4.1. A 21-version quantization test shows no quality difference: small models' bottleneck is base capability, not compression

May 52 min read