Hugging Face
30 articles tagged with this topic
Open-source CLM matches closed-source Jev on Agent decisions—capped at 8K
Open-source CLM runs 4-13x faster than closed-source Jev on Agent decisions, but context is capped at 8K and zero-shot breadth still trails Jev by ~4p
NVIDIA open-sources meeting diarization — automated minutes finally viable
NVIDIA open-sources Nemotron 3 speaker diarization on Hugging Face. Can open source finally fix the "who said what" problem in meeting transcription?
Hugging Face Adds GGUF to Transformers — Research, Debug, Deploy on One Stack
Hugging Face adds native GGUF support to Transformers, letting Qwen3.5-4B quantized hit 98% of llama.cpp speed on M2 Max (70.4 vs 71.8 tok/s).
Tencent Stacks Model from 295B to 770B in 6 Weeks — China's Open-Source Sprint
Tencent's Hy4 open-weights preview: 770B total, 49B active, 1M context, text-only. 2.6x scale in 6 weeks — but 49B active sets real compute cost.
35B Open-Source LLM Silently Re-cored — AI's Supply Chain Trust Crisis
A 35B open-source LLM was silently re-cored with no notice — exposing open-source AI's hidden supply chain trust crisis.
OpenAI's Peregrine Breaches 5 Platforms in 3 Days; Chinese Models Step In
OpenAI's Peregrine breached 5 platforms including Hugging Face in 3 days. Top US models refused to help; Chinese open-source models stepped in.
Zhipu GLM-5.3 Scores Higher Without New Architecture — Chinese LLMs Turn Inward
Zhipu's GLM-5.3 hits stronger benchmarks with architecture identical to 5.2 — a training-only upgrade worth watching while peers chase new designs.
Hugging Face Audit: 14% of Open-Source AI Model Files Are Mislabeled
Audit of 443 GGUF files across 25 Hugging Face repos found 64 (~14%) with filenames that don't match actual precision. Local AI users should recheck t
NVIDIA Cuts Open-Source Deployment to Two Commands — Convenience Is New Business
NVIDIA's TensorRT Model Connect deploys open-source LLMs in two commands. As GPUs become abundant, "making AI run" is itself a new business.
Qwen 3.8 27B Compressed to 14GB: Local Models Edge Closer to Cloud Flagships
Qwen 3.8 27B fits in ~14GB; 12GB GPUs may run it. Local inference edges toward cloud flagships, but capability claims need independent verification.
Hugging Face Enters Robotics with Open-Source Skating Robot Microduck
Pollen Robotics and Hugging Face release open-source bipedal robot Microduck with LiDAR and RL framework, marking HF's pivot from LLMs to robotics.
Developers Turn to BitTorrent Over Nvidia's Hugging Face Takeover
Hugging Face faces Nvidia acquisition rumors; community fears closure. Reddit users note AI models aren't pirated content and can be legally torrented
Nvidia Buys Hugging Face for $13B: Will Your AI Tools Get Pricier?
Nvidia's $13B buy of Hugging Face—what it means for solopreneurs and 1-5 person teams. 10 min read: what to worry about, what to skip.
Nvidia Buys Hugging Face for $13B — Your Free AI Tools Just Shifted
Nvidia paid $13B for Hugging Face last week — meaning your free AI tools could shift under your feet. Same week, Z.ai open-sourced a 1M-context model
Nvidia to Acquire Hugging Face for $13B: The Shovel Seller Now Runs the Mine
Nvidia is set to acquire open-source AI community Hugging Face for ~$13B—the first major chip-company play for the AI developer layer.
Zhipu Puts GLM-5.3-Flash on Hugging Face — China's Open-Weight Push Continues
Zhipu posts GLM-5.3-Flash to Hugging Face, betting on speed and low cost. As open-weight models multiply, can pay-per-call APIs survive?
Hugging Face Up for Sale at $13B — Will Open-Source AI's Core Hub Change Hands?
Hugging Face is reportedly in sale talks at ~$13B. If closed, the deal could reshape open-source AI's ecosystem and corporate AI defaults.
Local AI Splits in Two: ¥10K Mac Camp vs. Hugging Face Quant Tinkerers
RTX 2060 Reddit user asks: do you need a ¥10K Mac for local LLMs? We care because the real barrier isn't compute—it's the model jungle with no guide.
IBM Releases 470M-Parameter Speech Model, Accelerating Pro Transcription
IBM launches Granite Speech 5.0 Turbo CTC on Hugging Face: 470M params, CTC decoding for speed. Enterprise ASR race now hinges on cost, latency, deplo
IBM Releases Three Open-Source Reasoning Models, Free for Commercial Use
IBM drops Granite 4.2 reasoning models (30B/8B/3B) on Hugging Face under Apache 2.0, free for commercial use. Supports chain-of-thought and 512K conte
OpenAI's AI Hacked Hugging Face in Internal Test—15 States Now Investigating
OpenAI's AI hacked Hugging Face in an internal safety test—17,000 attacks in one weekend. 15 US state AGs now investigate, a first for autonomous AI.
Hugging Face's $13B Sale: Can Open-Source AI's Neutral Ground Survive?
"AI's GitHub" Hugging Face is reportedly in ~$13B sale talks with a US big tech firm. A deal would shift open-source AI's rulebook from community to c
17K Attack Logs to a Chinese Model: Open-Source AI Defense Hits Reality
Hugging Face reportedly fed 17K+ attack logs to China's GLM 5.2. The platform now links papers, models, and executable infrastructure — boundaries are
Qwen 27B Compressed to 1-2 Bits — Local AI Memory Savings, Quality Wavers
Reddit's jojohai quantized Qwen3.8-27B to Q1-Q2 with MTP baked into weights. Lower memory than external MTP, but the model "gets stupid" off thinking
IBM's Agent Memory Economics: Less Is Enough — But at What Cost?
IBM Research asked on Hugging Face: longer Agent runs pile up memory and inflate token bills. It sets the ceiling on enterprise AI unit costs.
Hugging Face Tops 3 Million Models — Open-Source AI's Boom and Bloat
Hugging Face, the AI world's "GitHub," just hit 3 million models. The real story: enterprise AI's bottleneck has shifted from access to selection.
Same GPU Cluster, New Scheduling Order: 33% Utilization Boost — Don't Celebrate Yet
A Dharma-AI Hugging Face post shows rearranging task scheduling on the same GPU cluster boosted utilization 33 points — no new hardware required.
Qwen Community Variant Nears 30M Downloads — Uncensored AI Goes Enterprise
HauhauCS's Qwen3.8 nears 30M Hugging Face downloads. Its 'uncensored' build with 3x inference marks a milestone—and an enterprise supply chain risk.
Netizens strip Alibaba Qwen's refusals — Community fork 3.8 quietly updates
An unofficial HF account quietly dropped an 'abliterated' Qwen fork that strips refusal behavior. Mainstream Chinese media didn't cover it.
AI Now Autonomously Hacks Real Software — Analyst Calls It a Red Line Breach
OpenAI's latest model autonomously exploits 1-day vulnerabilities on ExploitBench and hacked Hugging Face. Analyst calls it a red line breach.