AMD releases Instella-MoE-16B-A3B, open MoE LLM with 2.8B active params

AMD's Instella-MoE-16B-A3B is a fully open Mixture-of-Experts LLM with 16B total parameters and 2.8B active per token. It was trained from scratch on Instinct MI300X and MI325X GPUs, with AMD releasing weights from every training stage.
3 sources
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- 80 skills clone founder, philosopher, scientist minds in coding agents
- AI deployment startup June raises $20 million in pre-seed funding
- Project builds AI agents from scratch using local LLMs and node-llama-cpp
- Minimax + Sage Attention speed up video generation
- Trading Skills merges brokerage, charting, and screeners into Claude chat