Meta
26 articles tagged with this topic
Meta Open-Sources Million-GPU AI Network — Ditching Nvidia Goes Beyond Chips
Meta open-sources MetaRoCE—an RDMA protocol for million-GPU AI clusters. Not just tech: a crack in Nvidia's grip, this time at the network layer.
Meta Puts a 5.3B-Parameter Model Into 3GB of Phone Memory—For Research Only
Meta releases MobileMoE: 5.3B total parameters, under 1B active, and under 3GB at INT4—but its FAIR NC license bars commercial use.
Llama.cpp Hits 0.2 — The Bar for Running LLMs on Home PCs Drops Again
Open-source inference engine Llama.cpp releases 0.2.0, its first 0.2-series version, signaling a systemic architecture and performance overhaul worth
Nobody Trusts AI Benchmarks Anymore — Reddit Devs Call Out Score Inflation
A "which benchmark do you trust" thread on r/LocalLLaMA became a chorus of "I don't trust any." Our take: the entire AI benchmark system is failing.
Behind Gemma's 1 Billion Downloads: Google Is Serious on Open Source — but Late
Gemma downloads hit 1 billion; Google celebrates in SF with Demis Hassabis. Yet Meta and Alibaba have already seized the open-source throne.
Hyperscalers' Big Data Center Bet Wobbles on Cheaper Small Models
Hyperscalers' data center bets face repricing as investors push: AI's future may hinge on smaller models, not bigger ones.
White House First Targets Open-Source AI — Chinese LLMs' Overseas Play Reshuffles
Wired: White House updating AI policy framework — open-source LLMs in regulatory crosshairs. Meta, Alibaba, DeepSeek all in scope.
Training a 7B Model Takes 160GB VRAM — Why AI Is Now a Big-Players-Only Game
Training a 7B model demands 160GB VRAM — more than a single top-tier GPU can hold. This compute threshold makes AI a big-players-only game.
Meta's Muse Glimmer 30B Hits 3.3x Speedup on Mac as Open Source Closes the Gap
Developer A-Rahim uses speculative decoding to push Meta's Muse Glimmer 30B up to 3.3x faster on M4 Pro, with output exactly matching the original.
Three Open LLMs Dropped Same Day — The 'Frontier' Shelf Life Drops Below One Month
r/LocalLLaMA dubbed an ordinary Tuesday "Models Day" — at least three locally-runnable open-source LLMs landed in a single day. The gap between fronti
Meta Open-Sources 30B Agent — Your Life Data Is the Real Price
Meta open-sources its 30B Muse Glimmer Agent under Apache 2.0. The catch: it pitches "deep access to personal life" — a red flag for any ad business.
Zuckerberg personally drives Meta's release cadence — beating OpenAI's "polished" pace with speed
Zuckerberg reveals Meta's model release strategy this week: higher frequency, smaller steps. A direct challenge to OpenAI's "boutique slow-drop" appro
Meta Open-Sources 30B Local Agent Model — Big Tech Doubles Down on Putting AI Assistants on Your PC
Meta open-sources 30B-parameter Muse Glimmer, a local AI agent model hitting 20K tokens/sec on a single GPU, rewriting data privacy and cost calculus
Meta Crams AI Models Into Your Phone — But Is On-Device Really Worth It?
Meta open-sourced two phone-runnable models, Muse Glimmer and Muse Spark 1.2. We ask: is local inference actually cheaper than cloud APIs?
Meta Launches 30B Local Agent Model — Big Tech Finally Takes "AI On Your PC" Seriously
Meta releases open-source 30B Muse Glimmer, built for always-on local Agent workflows, fitting consumer GPUs when quantized. First major push to produ
Meta Buys Your Code for $0.2 — The Coding AI Battlefield Has Changed
Meta's Muse Code launches at $0.2/M tokens—if you let it train on your code. The coding AI race has shifted from "who's smarter" to "who's cheaper and
Your Content Is Training AI: What the Meta Lawsuit Exposed
Meta’s lawsuit reveals Big Tech is systematically harvesting your work. Spend 10 minutes adding one line to your config to slow down AI crawlers.
Meta ProgramBench: AI Still Can't Build Large Programs from Scratch
Meta ProgramBench tests AI building programs from scratch. Top models failed, cooling 'AI builds software' hype and exposing benchmark score inflation
1500 Bytes for Llama 2 Inference: Framework Bloat is a Choice, Not Inevitable
sectorllm achieves Llama 2 inference in <1500 bytes of x86 assembly. Core LLM logic is minimal; framework bloat is an engineering choice, not inevitab
Reddit's AI Hall of Fame: Giants Set the Tone, Community Does the Dirty Work
Reddit's open-source AI Hall of Fame covers Meta, DeepSeek, and llama.cpp. LLM prosperity depends on a strict community division of labor, not just bi
PyTorch Dominates 80% Dev Desktops—Nvidia Sells the Shovels in LLM Rush
PyTorch is the AI standard, but software unification exposes CUDA's hardware monopoly. LLM bottlenecks shifted from framework wars to GPU compute and
Meta Hardware Vault Locks Chat Backups: E2E Encryption Shifts to Storage
Meta updates WhatsApp/Messenger E2E encrypted backups. The encryption battle shifts from transit to storage, but transparency pledges lack binding pow
Meta's May 20 Lay offs Signal AI-Driven Workforce Restructuring Era
Meta's 10% global workforce reduction (~8,000 employees) explicitly tied to AI progress signals systematic labor displacement, forcing executives to
Meta's AI Agents Recover Hundreds of Megawatts by Automating Infrastructure Efficiency
Meta's unified AI agent platform compresses 10-hour manual regression investigations to 30 minutes, recovering hundreds of megawatts fleet -wide.
AI Neural Computer: Countdown to Software Disruption as Models Control Computers
Meta research enables AI to directly control terminal and desktop systems, threatening the 'humans operating software' employment model. SME owners mu
Meta Muse Spark: Hosted Model With 16 Built-In Chat Tools
Meta's first model since Llama 4, Muse Spark runs hosted-only with 16 exposed tools in meta .ai chat.