Not available in English yet
大模型窗口开到 128K 还是会'失忆' — 上下文工程师正成为抢手岗位
Related Reading
More on #DeepSeek
DeepSeek Hits 67 token/s on Two $9K Mini Boxes — Local LLM Floor Is Caving In
Reddit user hit 67-84 token/s on DeepSeek V4 Flash with a 1M-token context window on two ~$9K NVIDIA DGX Sparks. The local-LLM cost barrier is collaps
AI PPT Solutions Diverge Wildly — Two 67k-Star GitHub Projects Pick Sides
PPT Master (41k stars) outputs native Office; frontend-slides (26k stars) ships HTML. Two Claude Code Skills, opposite bets on AI's white-collar futur
OpenAI's Peregrine Breaches 5 Platforms in 3 Days; Chinese Models Step In
OpenAI's Peregrine breached 5 platforms including Hugging Face in 3 days. Top US models refused to help; Chinese open-source models stepped in.
Speculative decoding is becoming standard — open-source LLMs now predict ahead
Reddit users spotted speculative decoding working on local GPUs—AI instantly outputting phrases via MTP. The local inference cost curve is being quiet
Maxing AI 'Thinking Depth' Hurts Results — DGX Spark Local Test Warns Enterprises
A Reddit developer tested DeepSeek/Qwen on four DGX Sparks: 'deep thinking' mode lowers scores and doubles runtime — a direct cost warning for AI infe
4 Parallel Agents Beat 1: AI's Winning Play Shifts From Models to Systems
GPT-5.6 defaults to 4 parallel agents; NVIDIA's AVO aces ARC-AGI-3 — August signals say multi-agent is overtaking single-model scaling.