AMD releases Instella-MoE-16B-A3B, a fully open MoE LLM

The model packs 16B total parameters but activates only 2.8B per token, trained from scratch on AMD Instinct MI300X and MI325X GPUs. AMD is publishing weights from every training stage, along with data mixtures and training details.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Lap: open-source photo manager runs face recognition locally
- MemGraphRAG is a memory-based multi-agent system for Graph RAG
- MD-This-Page converts any webpage into LLM-ready Markdown
- NVIDIA releases Molt, a PyTorch-native agentic RL framework
- Superior Skills offers open-source agent trading schemas on Hyperliquid