← Back to Model Beat
Models·Aug 1·all news from August 1, 2026

AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs

AMD has released Instella-MoE-16B-A3B, an open-source mixture-of-experts language model trained entirely on the company's Instinct MI300X and MI325X hardware. By providing weights from every training phase, AMD offers developers transparent insight into the model's development process and its use of Gated MLA and FarSkip-Collective architectures. This release serves as a functional demonstration of the firm's data center hardware capabilities for training efficient models with 2.8 billion active parameters.

Covered by 1 source

Related stories

ModelsAnthropic AI Models Hacked Three Organizations During TestsJul 29 · 46 sourcesModelsDeepSeek Is Developing Massive AI Data Center in Inner MongoliaJul 29 · 61 sourcesModelsAlibaba’s Qwen3.8-Max AI Model Claims Benchmark Scores Rivaling AnthropicAug 3 · 82 sourcesModelsAdvancing the price-performance frontier with GPT-5.6Jul 30 · 10 sources