open source
17 articles tagged with this topic
llama.cpp Has 50 PRs Pending — Local AI No Longer Needs a High-End GPU
Open-source llama.cpp has 50+ performance PRs pending merge, some claiming 3x CPU inference speedup. Local LLM deployment is shedding its dependence o
Meta Puts a 5.3B-Parameter Model Into 3GB of Phone Memory—For Research Only
Meta releases MobileMoE: 5.3B total parameters, under 1B active, and under 3GB at INT4—but its FAIR NC license bars commercial use.
Qwen 27B Hits 138 Tokens/Sec on a Single RTX 3090 — Local AI Costs Crater
Qwen 27B hits 138 tokens/sec on one RTX 3090; 2nd-turn latency from 23s to 1s. Not a model breakthrough — open-source is flattening local LLM costs.
Unsloth squeezes Qwen onto 8GB laptops — local LLMs now run on almost anything
Unsloth's Dynamic v3.0 lets Qwen models run on 8GB laptops, with 1-bit versions keeping 77% accuracy — easing local enterprise LLM deployment.
Alibaba's Qwen Local Updates Again — Another Option for Running LLMs on Your PC
Alibaba's Qwen local GGUF build updates with lower VRAM needs. The bar for self-hosted LLMs drops a notch, but this is routine open-source maintenance
Local LLMs Get a GUI — But Hardware Costs Are Still the Wall
A GitHub project puts multiple local LLMs on Windows without code. The "data stays home" AI trend nears enterprise — but hardware costs bite.
DeepSeek Engineer: AI Streaming Is Three Channels, Not One Pipe
Developer traced DeepSeek's Harness via real logs. The popular "SSE long connection" story only covers one segment—three channels actually collaborate
DeepSeek Slices Agent Runtime Into Lego — China's LLM Firms Sell Foundations
DeepSeek open-sources Harness CLI: Agent runtime in pluggable Lego layers. China LLMs now sell foundations—developer preview, APIs may break.
AI Assistants Finally Remember You — And Memory Is Becoming a New Business
Headroom, an open-source project, lets AI assistants remember preferences across products and auto-learn from past mistakes. The AI race is shifting f
He Turned AI Into a "Virtual Software Company" — Multi-Agent Collaboration Moves from Concept to
An indie developer built TeamAgentX, an open source multi-agent framework that orchestrates Claude, Codex, and other coding agents like a real enginee
MIT Open-Source Tool Chamber: Cures AI's Citation Hallucination With Local Notes
A developer open-sourced Chamber under MIT license: local AI reads your notes, sees only numbered citations [1]…[k], not file paths, then resolves the
Tencent Open-Sources 3D World Generation — Runs on a Single 4090, but "Fun" Is Still Steps from
Tencent Hunyuan open-sources Hunyuan3D-WorldClaw: text-to-explorable 3D scenes on a single 4090. We assess why this AIGC tool is still a game-dev toy,
300B MoE on a 32GB Laptop — Local AI Finally Gets Usable
A developer ran DeepSeek V4's 300B MoE model on a 32GB laptop. The real story: hardware bottlenecks have shifted from compute to SSD read speed.
Alibaba Open-Sources Its Internal Code Review AI — 1/9 the Tokens, Exposing Generic Agents' Waste
Alibaba open-sources Open Code Review, its AI code review tool used by tens of thousands of engineers. The real finding: generic agents fail not at fi
有人开始用国产开源模型替换 Claude 做日常编程助手 — 性能差距正在缩小到「够用」
Developers on Reddit are seriously evaluating Alibaba's Qwen-35B-A3B as a local replacement for Claude Opus 4. 7 in daily coding workflows.
手机本地跑 AI 不再需要联网—— 一个开源安卓应用正在把这件事变得可操作
Pocket LLM v 1.4.0 shrinks to ~200MB, lets users download models on demand and run AI fully offline on Android.
Quoting Daniel Stenberg
cURL's lead dev reports AI -assisted security reports have improved in quality but surged in volume , consuming hours daily.