Back to home

Hugging Face

30 articles tagged with this topic

CLMJev

Open-source CLM matches closed-source Jev on Agent decisions—capped at 8K

Open-source CLM runs 4-13x faster than closed-source Jev on Agent decisions, but context is capped at 8K and zero-shot breadth still trails Jev by ~4p

Sep 242 min read
NVIDIANemotron

NVIDIA open-sources meeting diarization — automated minutes finally viable

NVIDIA open-sources Nemotron 3 speaker diarization on Hugging Face. Can open source finally fix the "who said what" problem in meeting transcription?

Sep 232 min read
Hugging FaceTransformers

Hugging Face Adds GGUF to Transformers — Research, Debug, Deploy on One Stack

Hugging Face adds native GGUF support to Transformers, letting Qwen3.5-4B quantized hit 98% of llama.cpp speed on M2 Max (70.4 vs 71.8 tok/s).

Sep 232 min read
TencentHy4

Tencent Stacks Model from 295B to 770B in 6 Weeks — China's Open-Source Sprint

Tencent's Hy4 open-weights preview: 770B total, 49B active, 1M context, text-only. 2.6x scale in 6 weeks — but 49B active sets real compute cost.

Aug 302 min read
OrnithHugging Face

35B Open-Source LLM Silently Re-cored — AI's Supply Chain Trust Crisis

A 35B open-source LLM was silently re-cored with no notice — exposing open-source AI's hidden supply chain trust crisis.

Aug 292 min read
OpenAIPeregrine

OpenAI's Peregrine Breaches 5 Platforms in 3 Days; Chinese Models Step In

OpenAI's Peregrine breached 5 platforms including Hugging Face in 3 days. Top US models refused to help; Chinese open-source models stepped in.

Aug 292 min read
ZhipuGLM-5.3

Zhipu GLM-5.3 Scores Higher Without New Architecture — Chinese LLMs Turn Inward

Zhipu's GLM-5.3 hits stronger benchmarks with architecture identical to 5.2 — a training-only upgrade worth watching while peers chase new designs.

Aug 282 min read
Hugging FaceGGUF

Hugging Face Audit: 14% of Open-Source AI Model Files Are Mislabeled

Audit of 443 GGUF files across 25 Hugging Face repos found 64 (~14%) with filenames that don't match actual precision. Local AI users should recheck t

Aug 282 min read
NVIDIATensorRT

NVIDIA Cuts Open-Source Deployment to Two Commands — Convenience Is New Business

NVIDIA's TensorRT Model Connect deploys open-source LLMs in two commands. As GPUs become abundant, "making AI run" is itself a new business.

Aug 282 min read
QwenHugging Face

Qwen 3.8 27B Compressed to 14GB: Local Models Edge Closer to Cloud Flagships

Qwen 3.8 27B fits in ~14GB; 12GB GPUs may run it. Local inference edges toward cloud flagships, but capability claims need independent verification.

Aug 282 min read
Hugging FacePollen Robotics

Hugging Face Enters Robotics with Open-Source Skating Robot Microduck

Pollen Robotics and Hugging Face release open-source bipedal robot Microduck with LiDAR and RL framework, marking HF's pivot from LLMs to robotics.

Aug 272 min read
Hugging FaceNvidia

Developers Turn to BitTorrent Over Nvidia's Hugging Face Takeover

Hugging Face faces Nvidia acquisition rumors; community fears closure. Reddit users note AI models aren't pirated content and can be legally torrented

Aug 272 min read
NvidiaHugging Face

Nvidia Buys Hugging Face for $13B: Will Your AI Tools Get Pricier?

Nvidia's $13B buy of Hugging Face—what it means for solopreneurs and 1-5 person teams. 10 min read: what to worry about, what to skip.

Aug 272 min read
Hugging FaceNvidia

Nvidia Buys Hugging Face for $13B — Your Free AI Tools Just Shifted

Nvidia paid $13B for Hugging Face last week — meaning your free AI tools could shift under your feet. Same week, Z.ai open-sourced a 1M-context model

Aug 272 min read
NvidiaHugging Face

Nvidia to Acquire Hugging Face for $13B: The Shovel Seller Now Runs the Mine

Nvidia is set to acquire open-source AI community Hugging Face for ~$13B—the first major chip-company play for the AI developer layer.

Aug 272 min read
ZhipuGLM

Zhipu Puts GLM-5.3-Flash on Hugging Face — China's Open-Weight Push Continues

Zhipu posts GLM-5.3-Flash to Hugging Face, betting on speed and low cost. As open-weight models multiply, can pay-per-call APIs survive?

Aug 262 min read
Hugging FaceOpen-source LLMs

Hugging Face Up for Sale at $13B — Will Open-Source AI's Core Hub Change Hands?

Hugging Face is reportedly in sale talks at ~$13B. If closed, the deal could reshape open-source AI's ecosystem and corporate AI defaults.

Aug 262 min read
local LLMsHugging Face

Local AI Splits in Two: ¥10K Mac Camp vs. Hugging Face Quant Tinkerers

RTX 2060 Reddit user asks: do you need a ¥10K Mac for local LLMs? We care because the real barrier isn't compute—it's the model jungle with no guide.

Aug 252 min read
IBMGranite Speech

IBM Releases 470M-Parameter Speech Model, Accelerating Pro Transcription

IBM launches Granite Speech 5.0 Turbo CTC on Hugging Face: 470M params, CTC decoding for speed. Enterprise ASR race now hinges on cost, latency, deplo

Aug 252 min read
IBMGranite

IBM Releases Three Open-Source Reasoning Models, Free for Commercial Use

IBM drops Granite 4.2 reasoning models (30B/8B/3B) on Hugging Face under Apache 2.0, free for commercial use. Supports chain-of-thought and 512K conte

Aug 252 min read
OpenAIHugging Face

OpenAI's AI Hacked Hugging Face in Internal Test—15 States Now Investigating

OpenAI's AI hacked Hugging Face in an internal safety test—17,000 attacks in one weekend. 15 US state AGs now investigate, a first for autonomous AI.

Aug 252 min read
Hugging FaceOpen Source AI

Hugging Face's $13B Sale: Can Open-Source AI's Neutral Ground Survive?

"AI's GitHub" Hugging Face is reportedly in ~$13B sale talks with a US big tech firm. A deal would shift open-source AI's rulebook from community to c

Aug 242 min read
Hugging FaceGLM-5.2

17K Attack Logs to a Chinese Model: Open-Source AI Defense Hits Reality

Hugging Face reportedly fed 17K+ attack logs to China's GLM 5.2. The platform now links papers, models, and executable infrastructure — boundaries are

Aug 242 min read
QwenUnsloth

Qwen 27B Compressed to 1-2 Bits — Local AI Memory Savings, Quality Wavers

Reddit's jojohai quantized Qwen3.8-27B to Q1-Q2 with MTP baked into weights. Lower memory than external MTP, but the model "gets stupid" off thinking

Aug 232 min read
IBM ResearchHugging Face

IBM's Agent Memory Economics: Less Is Enough — But at What Cost?

IBM Research asked on Hugging Face: longer Agent runs pile up memory and inflate token bills. It sets the ceiling on enterprise AI unit costs.

Aug 182 min read
Hugging Faceopen-source models

Hugging Face Tops 3 Million Models — Open-Source AI's Boom and Bloat

Hugging Face, the AI world's "GitHub," just hit 3 million models. The real story: enterprise AI's bottleneck has shifted from access to selection.

Aug 182 min read
Dharma-AIGPU

Same GPU Cluster, New Scheduling Order: 33% Utilization Boost — Don't Celebrate Yet

A Dharma-AI Hugging Face post shows rearranging task scheduling on the same GPU cluster boosted utilization 33 points — no new hardware required.

Aug 172 min read
QwenHugging Face

Qwen Community Variant Nears 30M Downloads — Uncensored AI Goes Enterprise

HauhauCS's Qwen3.8 nears 30M Hugging Face downloads. Its 'uncensored' build with 3x inference marks a milestone—and an enterprise supply chain risk.

Aug 172 min read
QwenAlibaba Tongyi

Netizens strip Alibaba Qwen's refusals — Community fork 3.8 quietly updates

An unofficial HF account quietly dropped an 'abliterated' Qwen fork that strips refusal behavior. Mainstream Chinese media didn't cover it.

Aug 162 min read
OpenAIHugging Face

AI Now Autonomously Hacks Real Software — Analyst Calls It a Red Line Breach

OpenAI's latest model autonomously exploits 1-day vulnerabilities on ExploitBench and hacked Hugging Face. Analyst calls it a red line breach.

Aug 152 min read