AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs
AMD unveiled Instella‑MoE‑16B‑A3B, a 16B parameter open‑source Mixture‑of‑Experts language model that activates only 2.8B parameters per token, trained from scratch on Instinct MI300X and MI325X GPUs.