Back to home
LLaMA
2 articles tagged with this topic
DeepSeekQwen
US-China LLMs Are Copying One Training Pipeline — Pretraining Isn't the Secret
US-China LLMs are converging on one training pipeline. Pretraining takes 90% of compute, but mid-training (5%) is what actually builds capability.
Aug 142 min read
Anubis-OSSApple Silicon
Local AI Gets Serious: Anubis-OSS Leaderboard Tracks 218 Models, 10 Apple Chips
Anubis-OSS leaderboard updates: 371 submissions, 218 models, 10 Apple chips. This data proves local open-source model deployment is no longer a geek t
May 52 min read