AMD Releases Fully Open-Source Instella-MoE Model with 16B Parameters on July 25

AMD-3.29%
According to BlockBeats, AMD released fully open-source mixture-of-experts (MoE) model Instella-MoE on July 25, featuring 16 billion total parameters with 2.8 billion parameters activated per token. The foundational model achieved an average score of 76.7, ranking among the leading open-source language models and outperforming SmolLM3-3B and OLMo-3-7B. Trained on AMD Instinct MI300X and MI325X GPUs using the ROCm software stack, Instella-MoE supports 64K token context length. AMD is publicly releasing all model weights, training configurations, data distributions, intermediate checkpoints, and inference code to advance open AI model research.
Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments