Apple
12 articles tagged with this topic
Two Mac Studios Hit 4.8TB/s — Home AI Takes On Data Centers, Community Skeptical
Exo Labs says two m5u Mac Studios hit 4.8TB/s memory bandwidth via RDMA. If real, local AI costs drop. Community is still verifying.
RTX 5090 Now Costs $5,090 — The Good Days of Running LLMs Locally Are Over
RTX 5090's street price hit $5,090, sparking despair on Reddit's local AI community. The consumer-GPU window for LLMs is closing — open-source local A
Qwen3.8 Hits 94% on Mac — But Benchmarks Are Breaking Down
Alibaba quietly released Qwen3.8; a developer hit 94% on a ~$7,000 Mac. More telling: the blogger admits "models are getting too good to differentiate
Mac mini's M6 upgrade drops the bar for running AI models locally
Apple's Mac mini refresh with next-gen Apple Silicon signals local AI is moving from geek toy toward semi-mainstream — but production readiness still
Mac Studio Packs 512GB RAM — Apple Takes Direct Aim at Cloud GPUs
Apple's new Mac Studio hits 512GB unified memory, finally making local 70B LLMs viable. Cloud GPU rental's monopoly cracks—but pricing locks out regul
Apple's M5 Ultra Hits 1.2TB/s Bandwidth — Local LLMs Cross the Practical Threshold
Apple's M5 Ultra hits 1.2TB/s memory bandwidth—the first chip making local LLM inference practical, reshaping how knowledge workers handle sensitive d
DeepSeek Model Hits 25 token/s on $7K Mac — Local AI Catches the Cloud
DeepSeek's latest model hits 25.8 tokens/s locally on an M2 Ultra Mac — smaller than official quant. Chinese open-source LLMs are now viable.
Mac Local LLM Inference: A Complete Mess, a Full Generation Behind NVIDIA
A Reddit developer tested all major Mac LLM frameworks for two weeks. Verdict: Apple's AI software is fragmented, a generation behind NVIDIA.
Local LLMs Finally 'Run' on Laptops — But Three Gaps Keep Them Off Your Work PC
Qwen 3.8 adds MTP (30–60% faster). Strix Halo and M4/M5 unified memory run 70B models on $2K laptops — but 'runnable' isn't 'work-ready.'
Your AI Wingman Today, Legal Freeze Tomorrow
It looks like Apple and OpenAI fighting over talent, but for small teams it’s really a reminder: don’t bet growth on one person or one platform.
Ollama Runs Local LLMs on Mac with One Command — PCs Are the New AI Gateway
Ollama runs Qwen & DeepSeek locally on Mac via one command. MLX integration doubles inference speed. When deployment = app install, cloud-free AI may
Apple Foldable iPhone Faces Engineering Delays, Sources Say
Reports indicate Apple's foldable iPhone has hit engineering obstacles that could push back its shipping timeline.