Back to home

open source

17 articles tagged with this topic

llama.cppMoE

llama.cpp Has 50 PRs Pending — Local AI No Longer Needs a High-End GPU

Open-source llama.cpp has 50+ performance PRs pending merge, some claiming 3x CPU inference speedup. Local LLM deployment is shedding its dependence o

7h ago2 min read
MetaMobileMoE

Meta Puts a 5.3B-Parameter Model Into 3GB of Phone Memory—For Research Only

Meta releases MobileMoE: 5.3B total parameters, under 1B active, and under 3GB at INT4—but its FAIR NC license bars commercial use.

5d ago2 min read
QwenvLLM

Qwen 27B Hits 138 Tokens/Sec on a Single RTX 3090 — Local AI Costs Crater

Qwen 27B hits 138 tokens/sec on one RTX 3090; 2nd-turn latency from 23s to 1s. Not a model breakthrough — open-source is flattening local LLM costs.

Aug 192 min read
UnslothQwen

Unsloth squeezes Qwen onto 8GB laptops — local LLMs now run on almost anything

Unsloth's Dynamic v3.0 lets Qwen models run on 8GB laptops, with 1-bit versions keeping 77% accuracy — easing local enterprise LLM deployment.

Aug 192 min read
AlibabaQwen

Alibaba's Qwen Local Updates Again — Another Option for Running LLMs on Your PC

Alibaba's Qwen local GGUF build updates with lower VRAM needs. The bar for self-hosted LLMs drops a notch, but this is routine open-source maintenance

Aug 192 min read
llama.cpplocal AI

Local LLMs Get a GUI — But Hardware Costs Are Still the Wall

A GitHub project puts multiple local LLMs on Windows without code. The "data stays home" AI trend nears enterprise — but hardware costs bite.

Aug 162 min read
DeepSeekHarness

DeepSeek Engineer: AI Streaming Is Three Channels, Not One Pipe

Developer traced DeepSeek's Harness via real logs. The popular "SSE long connection" story only covers one segment—three channels actually collaborate

Aug 152 min read
DeepSeekHarness

DeepSeek Slices Agent Runtime Into Lego — China's LLM Firms Sell Foundations

DeepSeek open-sources Harness CLI: Agent runtime in pluggable Lego layers. China LLMs now sell foundations—developer preview, APIs may break.

Aug 152 min read
HeadroomClaude

AI Assistants Finally Remember You — And Memory Is Becoming a New Business

Headroom, an open-source project, lets AI assistants remember preferences across products and auto-learn from past mistakes. The AI race is shifting f

Aug 142 min read
TeamAgentXmulti-agent

He Turned AI Into a "Virtual Software Company" — Multi-Agent Collaboration Moves from Concept to

An indie developer built TeamAgentX, an open source multi-agent framework that orchestrates Claude, Codex, and other coding agents like a real enginee

Aug 102 min read
ChamberMIT

MIT Open-Source Tool Chamber: Cures AI's Citation Hallucination With Local Notes

A developer open-sourced Chamber under MIT license: local AI reads your notes, sees only numbered citations [1]…[k], not file paths, then resolves the

Aug 92 min read
TencentHunyuan

Tencent Open-Sources 3D World Generation — Runs on a Single 4090, but "Fun" Is Still Steps from

Tencent Hunyuan open-sources Hunyuan3D-WorldClaw: text-to-explorable 3D scenes on a single 4090. We assess why this AIGC tool is still a game-dev toy,

Aug 92 min read
DeepSeekMoE

300B MoE on a 32GB Laptop — Local AI Finally Gets Usable

A developer ran DeepSeek V4's 300B MoE model on a 32GB laptop. The real story: hardware bottlenecks have shifted from compute to SSD read speed.

Aug 92 min read
AlibabaOpen Code Review

Alibaba Open-Sources Its Internal Code Review AI — 1/9 the Tokens, Exposing Generic Agents' Waste

Alibaba open-sources Open Code Review, its AI code review tool used by tens of thousands of engineers. The real finding: generic agents fail not at fi

Aug 92 min read
QwenClaude

有人开始用国产开源模型替换 Claude 做日常编程助手 — 性能差距正在缩小到「够用」

Developers on Reddit are seriously evaluating Alibaba's Qwen-35B-A3B as a local replacement for Claude Opus 4. 7 in daily coding workflows.

Apr 202 min read
Pocket LLMon-device AI

手机本地跑 AI 不再需要联网—— 一个开源安卓应用正在把这件事变得可操作

Pocket LLM v 1.4.0 shrinks to ~200MB, lets users download models on demand and run AI fully offline on Android.

Apr 192 min read
cURLDaniel Stenberg

Quoting Daniel Stenberg

cURL's lead dev reports AI -assisted security reports have improved in quality but surged in volume , consuming hours daily.

Apr 92 min read