August 1, 2026 · MarkTechPost
AMD releases fully open MoE model trained on its own Instinct GPUs
AMD released Instella-MoE-16B-A3B, an open mixture-of-experts language model with 16 billion total parameters but only 2.8 billion active per token, trained from scratch on Instinct MI300X and MI325X GPUs. AMD published the full training pipeline -- weights from every stage, data mixtures, configs, and inference code.
Why it matters: This is as much a proof point for AMD's Instinct GPU line as a model release -- training a full LLM end-to-end on AMD hardware and open-sourcing the entire pipeline gives outside developers a reference for building on non-Nvidia infrastructure. That matters for reducing the industry's dependence on Nvidia for both training and open-model reproducibility.