Not available in English yet
Google 开源 Gemma 4 长文本翻车 — 本地跑大模型的坑比想象中深
Related Reading
More on #Google
AI at 1/4 Size, 96% Capability — Local AI Cost Tipping Point Is Here
A 3.3GB small model jumped from 28.9 to 69.5 on reasoning via precision allocation — usable local AI may cost less than we thought.
Google Doubles Gemma 4 Speed — Speculative Decoding Goes Mainstream
Google's Gemma 4 MTP models use speculative decoding for up to 2x speed with zero quality loss, boosting local LLM practicality and lowering compute b
Google Gemma 4 Fixes Chat Template — Local LLM Usability Inches Forward
Google fixed Gemma 4's chat template bug; community quantized versions updated. Not major news, but proves local AI usability inches up via detail ref
Google Open-Sources 'K8s for Agents' AX — 11.6K Stars but Still Alpha
Google open-sources AX, its 'K8s for Agents': declarative YAML scheduling, 11.6K Stars—but Alpha with documented breaking-change risks.
95 Tokens for Cloud Plans, 97% Stay Local: Task Routing Outweighs the Model
Antigravity's local-model support shows task routing in action: cloud gets summaries, local runs execution. Beats bigger models on enterprise complian
Google Gemma Developers Demand Updates on Reddit — Open-Source LLMs Lag Behind
Google's open-source Gemma LLM faces Reddit backlash—users demand a 220B MoE version as open vs closed-source competition intensifies.