AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs
Article summary
Quick briefing — cleaned from the original RSS feed
AMD released Instella-MoE-16B-A3B, a fully open Mixture-of-Experts language model trained from scratch on Instinct MI300X and MI325X GPUs. It holds 16B total parameters but activates only 2.8B per token, using Gated MLA and FarSkip-Collective. AMD published weights from every training stage, plus data mixtures, configs, and inference code.
1Key Takeaways
- AMD released Instella-MoE-16B-A3B, a fully open Mixture-of-Experts language model trained from scratch on Instinct MI300X and MI325X GPUs.
- It holds 16B total parameters but activates only 2.8B per token, using Gated MLA and FarSkip-Collective.
- AMD published weights from every training stage, plus data mixtures, configs, and inference code.
2AIWedia Score
8.1/10
High relevance — worth your attention today
Based on source trust, recency, category impact, and story depth.
3Why it matters
Video AI is reshaping ads, social content, and entertainment with faster generation pipelines. MarkTechPost Video reports that aMD released Instella-MoE-16B-A3B, a fully open Mixture-of-Experts language model trained from scratch on Instinct MI300X and MI325X GPUs.
Explore related
Browse toolsVideo AI news
Explore curated video ai tools on AIWedia — compare, rank, and launch from our directory.
Full story on MarkTechPost Video
Read full articleHeadlines aggregated via RSS for discovery on AIWedia. Original content © MarkTechPost Video. We link to the source and do not republish full articles.