Back to home
edge AI
5 articles tagged with this topic
QwenAlibaba
Qwen 1-bit is still 6x slower — companies eyeing local LLMs should wait
Qwen's 1-bit quantization runs 6x slower than the 4-bit version with ~70% accuracy. Local LLM deployment isn't ready to replace cloud APIs.
Just now2 min read
audio.cppopen-source
Local AI audio models balloon to 62 — an arena helps you choose
audio.cpp 0.7 ships 62 model families and 85 variants, plus an Arena comparison interface — selection tools signal maturing open-source voice AI.
2d ago2 min read
HorizonJ6
Horizon J6 Runs YOLOv5 End-to-End: How Far to Plug-and-Play for Chinese AI Chips?
YOLOv5 runs end-to-end on Horizon J6: conversion, C++ inference, accuracy. Chinese AI chips advance on vision, but engineering stays hard.
Aug 202 min read
Liquid AILFM2.5-VL-3B
Liquid AI packs 2GB vision model on phones — local AI finally tells what, where
Liquid AI ships LFM2.5-VL-3B: 2GB, 3.1B params, runs locally on iPhone 17. ScreenSpot-v2 score jumped 6 to 78.7 — local models now truly see.
Aug 122 min read
VLX-Seekomlab
Korean team swaps bounding boxes for region selection — robot local vision models get lighter
Korean open-source team omlab releases 10B VLX-Seek 1.5: instead of outputting coordinates, the model selects candidate regions for grounding. A stead
Aug 102 min read