Back to home

open-source models

30 articles tagged with this topic

Qwenlocal LLM

Local AI Coding Is Trending — But Most Companies' GPUs Can't Run It

A Reddit post about running Qwen 3 27B locally on an RTX A4500 for AI coding sparked debate. Local model coding is shifting from hobbyist toy to real

1h ago2 min read
OrnithHugging Face

35B Open-Source LLM Silently Re-cored — AI's Supply Chain Trust Crisis

A 35B open-source LLM was silently re-cored with no notice — exposing open-source AI's hidden supply chain trust crisis.

9h ago2 min read
QwenAlibaba

Alibaba Open-Source Model Revives a Bricked Foldable — Local AI Runs Solo

Reddit user revived a bricked foldable with Qwen 3.8 (27B) on a Raspberry Pi, saving ~$600. Real signal: local AI trusted solo on critical tasks.

19h ago2 min read
RTX 5090Nvidia

RTX 5090 Now Costs $5,090 — The Good Days of Running LLMs Locally Are Over

RTX 5090's street price hit $5,090, sparking despair on Reddit's local AI community. The consumer-GPU window for LLMs is closing — open-source local A

2d ago2 min read
QwenAlibaba

Alibaba's Qwen Tops Both Open-Source Frontiers — Parameters Weren't the Question

Alibaba's Qwen hits the Pareto frontier on both total and active parameters. Running AI may keep getting cheaper — but the lead may not hold.

2d ago2 min read
QwenDeepSeek

User Claims Qwen Beats DeepSeek — But the Post Has Zero Content

A Reddit post claims Qwen3.8-Flash-Next beats DeepSeek V4 Pro-no benchmarks, no tables. We unpack the anxiety behind this empty post.

2d ago2 min read
ZhipuQwen

Zhipu GLM Flash Beats Qwen — Size Doesn't Cut It

GLM Flash beat Qwen on three of four Reddit benchmarks. Qwen only led on graduate-level science by half a point — despite GLM being the larger model.

3d ago2 min read
QwenNVIDIA

Lunchbox Rig Runs 27B AI Model — Local Players Catch Up to Paid Services

A Reddit rig pairs a Toughbook with an RTX Pro 6000 to run a 27B model at 262K context — and beats Gemini Pro and ChatGPT on legal OCR.

4d ago2 min read
QwenNVIDIA

Local AI Is Finally Competitive — What 30B Open-Source Benchmarks Reveal

LocalLLaMA benchmark shows 30B open-source models (Ornith, TielCoder, Qwen, Nemotron) now rival closed APIs on coding—local AI deployment becomes viab

5d ago2 min read
QwenTongyi Qianwen

Qwen Autonomously Writes a C Compiler — Open-Source Agents Enter the Long-Haul Race

Qwen 3.6 27B built a C99 compiler autonomously in 6 weeks. Open-source handles long-haul tasks, but hallucinations and context issues still put "auton

5d ago2 min read
QwenTongyi Qianwen

Qwen 27B Lets Local AI Match Gemini — China's Small-Model Reversal

Qwen 27B matches or beats Gemini on OCR and code, per Reddit. A US team eyes self-hosted AI with under-two-month payback. Local AI feels usable.

6d ago2 min read
QwenAlibaba

Qwen 3.8-27B One Week In: Top Marks for Doing, Memory Slips

Qwen 3.8-27B earned 'local best' on agent tasks in 2,000 Reddit tests but regressed on knowledge memory—first open-source model approaching GPT on exe

6d ago2 min read
h3.cApple Silicon

33B Audio-Video Model on Apple Silicon: Open Source Tallies Acceleration's True Cost

This week, h3.c ported a 33B audio-video model to Apple Silicon, exposing six distinct optimization knobs instead of a single "fast=true" switch.

Aug 222 min read
AlibabaQwen

Alibaba Qwen Ships Twice in 6 Months — Local-Usable Builds Always Lag 2 Months

Alibaba's Qwen shipped two versions in six months, but the locally deployable versions always trail the benchmark leader by two months.

Aug 222 min read
QwenAlibaba

Qwen 27B Adds "Thinking Levels" — Small Models Catch Up via Efficiency, Not Size

Alibaba's Qwen3.8-27B now supports adjustable thinking levels. Reddit testers say even the lowest tier beats the previous generation — a sign mid-size

Aug 212 min read
dots3-notellama.cpp

dots3-note Goes Open Source: 280B Params, 512K Context, AI Note-Taking Bet

dots3-note entered llama.cpp this week: 280B params, 512K context, multimodal, focused on long-context memory and note-style learning.

Aug 212 min read
AQuALocalLLaMA

LocalLLaMA Debunks 'AI Self-Improvement': Weights Untouched, Notes Upgraded

AQuA preprint debunked: 'AI self-improvement' is just state updates, weights untouched. An old reproducibility problem every AI buyer should care abou

Aug 202 min read
QwenTongyi Qwen

Qwen 27B on RTX 3090s draws nested SVG; mid-size open-source models underrated

Qwen 27B drew nested SVG on two RTX 3090s at 46k tokens. Mid-size open-source creativity on consumer hardware looks underrated; ~$10K GPUs stay a barr

Aug 202 min read
QwenAlibaba

Open-source 27B model proves 'slow thinking' — test-time compute hits mid-tier

Reddit's local LLM community drops a Qwen-based 27B variant: slower per-token, better quality. First 'slow thinking' win for mid-sized open models.

Aug 182 min read
Hugging Faceopen-source models

Hugging Face Tops 3 Million Models — Open-Source AI's Boom and Bloat

Hugging Face, the AI world's "GitHub," just hit 3 million models. The real story: enterprise AI's bottleneck has shifted from access to selection.

Aug 182 min read
QwenAlibaba

Qwen Thinks 90 Min, No Answer — One Parameter Cures Open-Source Overthinking

Alibaba's Qwen3.8-27b overthinks by default — inferences can run 90 minutes without output. A Reddit dev shared two llamacpp params that fix it.

Aug 172 min read
QwenAlibaba Tongyi Qianwen

Qwen 3.8 Open-Source: 27B on RTX 4090 — Closed-Source Flagship Moat Loosens

Qwen 3.8 open-weights: 27B beats Claude Opus 4.6 Max on coding/agent benchmarks. 4-bit fits 24GB. New option for enterprise IT, dev tools, hardware.

Aug 172 min read
AlibabaQwen3

Alibaba's Qwen3 27B Compressed to 18GB, Runs on Single GPU — Local LLM Bar Drops Again

Alibaba's Qwen3 27B compressed to 18GB via int4 quantization with MTP acceleration, runs on a single consumer GPU. Local LLM hardware costs keep falli

Aug 162 min read
Qwenopen-source models

Local Open-Source LLMs Booming for Coding — The Real Battle Is Now the Harness

Reddit poll this week: developers now pick the harness that wires AI into their code workflow. The fight is moving from the model to the tool layer.

Aug 162 min read
QwenMuse Glimmer

Qwen Loses to Muse Glimmer in Image Recognition: Is the Moat Cracking?

LocalLLaMA test: Muse Glimmer 30B swept Qwen 3.8 27B on OCR and image-text tasks. Informal, but exposes Qwen's multimodal weakness.

Aug 152 min read
AlibabaQwen

Alibaba's Qwen 27B: Three Iterations, A Visible One-Shot Lift

Qwen 27B across three versions: 2.46 → 3.00 on 35 one-shot tasks. Small gains, stable direction. Worth watching for open-source LLM watchers.

Aug 152 min read
QwenAlibaba

Alibaba Qwen Drops 27B Open-Source — On-Prem AI Is Finally Within Reach

Qwen 27B is now open-source, free for commercial use. The community built consumer-grade versions within 24 hours—27B hits the capability-cost sweet s

Aug 152 min read
Liquid AILFM

Liquid AI Drops Cookbook Tutorials — Small Model Vendors Court Developers

Liquid AI dropped a developer cookbook on GitHub this week. Our interest: small-model vendors, outgunned on parameter scale, now compete on developer

Aug 122 min read
GoogleDeepMind

Google DeepMind's WeatherNext 2 Adds a Full Day to Typhoon Forecasts — AI Is Becoming Weather

Google DeepMind's open-source WeatherNext 2 model extends typhoon forecast lead time from 2 to 3 days with higher accuracy, signaling AI is replacing

Aug 92 min read
DeepSeekV4-Flash

DeepSeek V4-Flash Falls Asleep Mid-Task: The Hidden Cost of Local AI Deployment

DeepSeek's V4-Flash open-source model silently halts generation past 100k tokens. Concrete proof that "downloadable" and "reliably usable" are still f

Aug 92 min read