Back to home

HuggingFace

14 articles tagged with this topic

UnslothHuggingFace

HF 收购疑云下,Unsloth 让普通显卡跑得动大模型 — 开源生态的真正护城河

Under HF acquisition rumors, Unsloth lets 24GB GPUs run 70B models. Small open-source teams—not platforms—will decide if AI stays affordable.

2d ago2 min read
NvidiaHuggingFace

Nvidia 吞下 HuggingFace,本地跑 AI 的核心工具被收编 — 开源社区最怕的事正在发生

Nvidia's acquisition of HuggingFace pulls in the llama.cpp team and copyright—the de facto standard for running LLMs locally. Community fears license

2d ago2 min read
HuggingFaceNvidia

BT instead of HuggingFace? Reddit user pitches it — three big hurdles

50GB+ open-source models strain HuggingFace's bandwidth budget. Reddit's BT-sharing idea sounds win-win — but security, dead torrents, and business vi

2d ago2 min read
benchmarkllm-evaluation

Nobody Trusts AI Benchmarks Anymore — Reddit Devs Call Out Score Inflation

A "which benchmark do you trust" thread on r/LocalLLaMA became a chorus of "I don't trust any." Our take: the entire AI benchmark system is failing.

Aug 222 min read
OrnithMTP

Ornith 1.5 Shipped with an Untrained MTP Head — Open-Source AI's QC Problem

Ornith 1.5 shipped with an untrained MTP head. Not a bug — a symptom of QA gaps in open-source AI that any cost-cutting enterprise should heed.

Aug 202 min read
Aurora-80KLocalLLaMA

80K-parameter AI model drops — not a ChatGPT rival, but a signal

Independent dev releases Aurora-80K, an 80K-parameter open-source language model. Far below mainstream LLMs, but it shows how small AI models can get.

Aug 202 min read
DeepSeekV4-Pro

DeepSeek V4 Pro Reignites Open vs Closed-Source Battle in AI

DeepSeek released V4 Pro 0813 open source on HuggingFace. Two Minute Papers titled it "making closed AI look ridiculous." We unpack the open vs closed

Aug 192 min read
QwenHuggingFace

Qwen Hit 1M Downloads, Under 1,000 Can Run It: Local LLMs Stay a Geek Toy

A Reddit user estimated Qwen 3 27B has ~1M HuggingFace downloads, but fewer than 1,000 people globally own 24GB+ GPUs to actually run it locally.

Aug 162 min read
MiniMaxH3

MiniMax H3 Open-Source Video Model Hits 2K — Local Deployment Barrier Drops a Notch

MiniMax released the full weights of its H3 multimodal video model in early August: 2K, 24fps, native stereo, open-source and locally runnable. The sa

Aug 92 min read
AI PolicyOpen Source AI

Gov AI Veto: How Solo Founders Prep

US AI model reviews might leave small teams and open-source last to access top tools. Diversify dependencies early and avoid getting stuck.

May 62 min read
RedditMeta

Reddit's AI Hall of Fame: Giants Set the Tone, Community Does the Dirty Work

Reddit's open-source AI Hall of Fame covers Meta, DeepSeek, and llama.cpp. LLM prosperity depends on a strict community division of labor, not just bi

May 32 min read
QwenSAE

Qwen Open-Sources SAE: Decoding & Steering LLMs, China Enters Interpretability

Qwen open-sourced an 80K-feature SAE on HuggingFace. For the first time, a Chinese team makes LLM internals dissectible & steerable—a major interpreta

May 32 min read
GemmaGoogle

Gemma 4 Hits HuggingFace — Open Source Outpaces Official Toolchain

gemma-4-31B-it-DFlash on HuggingFace lacks llama.cpp support. We see models outpacing toolchains—having models you can't run is the new paradox.

May 22 min read
QwenQwen3.6- 35B-A3B

Qwen3.6-35B-A3B released!

Alibaba's Qwen team releases a 35B sparse MoE model with only 3B active params under Apache 2.0.

Apr 162 min read