Back to home

Meta

26 articles tagged with this topic

MetaMetaRoCE

Meta Open-Sources Million-GPU AI Network — Ditching Nvidia Goes Beyond Chips

Meta open-sources MetaRoCE—an RDMA protocol for million-GPU AI clusters. Not just tech: a crack in Nvidia's grip, this time at the network layer.

Aug 242 min read
MetaMobileMoE

Meta Puts a 5.3B-Parameter Model Into 3GB of Phone Memory—For Research Only

Meta releases MobileMoE: 5.3B total parameters, under 1B active, and under 3GB at INT4—but its FAIR NC license bars commercial use.

Aug 242 min read
Llama.cppMeta

Llama.cpp Hits 0.2 — The Bar for Running LLMs on Home PCs Drops Again

Open-source inference engine Llama.cpp releases 0.2.0, its first 0.2-series version, signaling a systemic architecture and performance overhaul worth

Aug 222 min read
benchmarkllm-evaluation

Nobody Trusts AI Benchmarks Anymore — Reddit Devs Call Out Score Inflation

A "which benchmark do you trust" thread on r/LocalLLaMA became a chorus of "I don't trust any." Our take: the entire AI benchmark system is failing.

Aug 222 min read
GoogleGemma

Behind Gemma's 1 Billion Downloads: Google Is Serious on Open Source — but Late

Gemma downloads hit 1 billion; Google celebrates in SF with Demis Hassabis. Yet Meta and Alibaba have already seized the open-source throne.

Aug 202 min read
MicrosoftGoogle

Hyperscalers' Big Data Center Bet Wobbles on Cheaper Small Models

Hyperscalers' data center bets face repricing as investors push: AI's future may hinge on smaller models, not bigger ones.

Aug 202 min read
White HouseOpen-Source LLMs

White House First Targets Open-Source AI — Chinese LLMs' Overseas Play Reshuffles

Wired: White House updating AI policy framework — open-source LLMs in regulatory crosshairs. Meta, Alibaba, DeepSeek all in scope.

Aug 132 min read
DeepSpeedMicrosoft

Training a 7B Model Takes 160GB VRAM — Why AI Is Now a Big-Players-Only Game

Training a 7B model demands 160GB VRAM — more than a single top-tier GPU can hold. This compute threshold makes AI a big-players-only game.

Aug 132 min read
MetaMuse Glimmer

Meta's Muse Glimmer 30B Hits 3.3x Speedup on Mac as Open Source Closes the Gap

Developer A-Rahim uses speculative decoding to push Meta's Muse Glimmer 30B up to 3.3x faster on M4 Pro, with output exactly matching the original.

Aug 122 min read
r/LocalLLaMAOpen-source LLMs

Three Open LLMs Dropped Same Day — The 'Frontier' Shelf Life Drops Below One Month

r/LocalLLaMA dubbed an ordinary Tuesday "Models Day" — at least three locally-runnable open-source LLMs landed in a single day. The gap between fronti

Aug 122 min read
MetaMuse Glimmer

Meta Open-Sources 30B Agent — Your Life Data Is the Real Price

Meta open-sources its 30B Muse Glimmer Agent under Apache 2.0. The catch: it pitches "deep access to personal life" — a red flag for any ad business.

Aug 122 min read
MetaMark Zuckerberg

Zuckerberg personally drives Meta's release cadence — beating OpenAI's "polished" pace with speed

Zuckerberg reveals Meta's model release strategy this week: higher frequency, smaller steps. A direct challenge to OpenAI's "boutique slow-drop" appro

Aug 102 min read
MetaMuse Glimmer

Meta Open-Sources 30B Local Agent Model — Big Tech Doubles Down on Putting AI Assistants on Your PC

Meta open-sources 30B-parameter Muse Glimmer, a local AI agent model hitting 20K tokens/sec on a single GPU, rewriting data privacy and cost calculus

Aug 102 min read
MetaMuse Glimmer

Meta Crams AI Models Into Your Phone — But Is On-Device Really Worth It?

Meta open-sourced two phone-runnable models, Muse Glimmer and Muse Spark 1.2. We ask: is local inference actually cheaper than cloud APIs?

Aug 102 min read
MetaMuse Glimmer

Meta Launches 30B Local Agent Model — Big Tech Finally Takes "AI On Your PC" Seriously

Meta releases open-source 30B Muse Glimmer, built for always-on local Agent workflows, fitting consumer GPUs when quantized. First major push to produ

Aug 102 min read
MetaMuse Code

Meta Buys Your Code for $0.2 — The Coding AI Battlefield Has Changed

Meta's Muse Code launches at $0.2/M tokens—if you let it train on your code. The coding AI race has shifted from "who's smarter" to "who's cheaper and

Aug 92 min read
MetaCopyright Protection

Your Content Is Training AI: What the Meta Lawsuit Exposed

Meta’s lawsuit reveals Big Tech is systematically harvesting your work. Spend 10 minutes adding one line to your config to slow down AI crawlers.

May 62 min read
MetaProgramBench

Meta ProgramBench: AI Still Can't Build Large Programs from Scratch

Meta ProgramBench tests AI building programs from scratch. Top models failed, cooling 'AI builds software' hype and exposing benchmark score inflation

May 62 min read
sectorllmLlama2

1500 Bytes for Llama 2 Inference: Framework Bloat is a Choice, Not Inevitable

sectorllm achieves Llama 2 inference in <1500 bytes of x86 assembly. Core LLM logic is minimal; framework bloat is an engineering choice, not inevitab

May 52 min read
RedditMeta

Reddit's AI Hall of Fame: Giants Set the Tone, Community Does the Dirty Work

Reddit's open-source AI Hall of Fame covers Meta, DeepSeek, and llama.cpp. LLM prosperity depends on a strict community division of labor, not just bi

May 32 min read
PyTorchNvidia

PyTorch Dominates 80% Dev Desktops—Nvidia Sells the Shovels in LLM Rush

PyTorch is the AI standard, but software unification exposes CUDA's hardware monopoly. LLM bottlenecks shifted from framework wars to GPU compute and

May 32 min read
MetaWhatsApp

Meta Hardware Vault Locks Chat Backups: E2E Encryption Shifts to Storage

Meta updates WhatsApp/Messenger E2E encrypted backups. The encryption battle shifts from transit to storage, but transparency pledges lack binding pow

May 12 min read
layoffsMeta

Meta's May 20 Lay offs Signal AI-Driven Workforce Restructuring Era

Meta's 10% global workforce reduction (~8,000 employees) explicitly tied to AI progress signals systematic labor displacement, forcing executives to

Apr 182 min read
MetaFBDetect

Meta's AI Agents Recover Hundreds of Megawatts by Automating Infrastructure Efficiency

Meta's unified AI agent platform compresses 10-hour manual regression investigations to 30 minutes, recovering hundreds of megawatts fleet -wide.

Apr 162 min read
AI automationMeta

AI Neural Computer: Countdown to Software Disruption as Models Control Computers

Meta research enables AI to directly control terminal and desktop systems, threatening the 'humans operating software' employment model. SME owners mu

Apr 122 min read
MetaMuse Spark

Meta Muse Spark: Hosted Model With 16 Built-In Chat Tools

Meta's first model since Llama 4, Muse Spark runs hosted-only with 16 exposed tools in meta .ai chat.

Apr 92 min read