Open SourceAug 1, 2026, 7:01 PM

AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs

30-second summary

AMD released Instella-MoE-16B-A3B, a fully open Mixture-of-Experts LLM with 16B total parameters but only 2.8B active per token. The model was trained on Instinct MI300X and MI325X GPUs and includes full training weights, data mixtures, and inference code.

TickrWire
AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs
Key takeaways
  • Instella-MoE-16B-A3B is a fully open Mixture-of-Experts LLM with 16B total parameters but only 2.8B active per token, reducing computational load.
  • The model was trained on AMD's Instinct MI300X and MI325X GPUs and includes full training weights, data mixtures, and inference code.
  • AMD leverages Gated MLA and FarSkip-Collective techniques to enhance efficiency and performance in the MoE architecture.
  • This release highlights AMD's commitment to open AI models and its growing presence in the AI hardware and software market.
Full story

AMD has introduced Instella-MoE-16B-A3B, a fully open Mixture-of-Experts (MoE) language model featuring 16 billion total parameters but activating just 2.8 billion per token. This design leverages Gated MLA and FarSkip-Collective techniques to optimize efficiency while maintaining performance. The model was trained from scratch on AMD's Instinct MI300X and MI325X GPUs, positioning it as a high-performance option for developers seeking open alternatives in the MoE space.

The release includes comprehensive resources: AMD has published the model weights from every training stage, along with the data mixtures used, configuration files, and inference code. This transparency aims to foster community adoption and further research, aligning with the growing demand for open and reproducible AI models. The move also underscores AMD's push into the AI hardware and software ecosystem, competing directly with established players in the LLM market.

Sponsored
Why this matters
Developers

Provides a fully open, efficient MoE model with transparent training resources, enabling customization and research.

Businesses

Offers a competitive open alternative to proprietary MoE models, potentially reducing costs and dependency on closed solutions.

Investors

Signals AMD's strategic expansion into AI hardware and software, with potential long-term market impact.

Everyone

Demonstrates the growing trend of open, efficient AI models and AMD's role in shaping the future of AI infrastructure.

Glossary
Mixture-of-Experts (MoE)
A neural network architecture where multiple specialized sub-models (experts) are combined, with only a subset activated per input to improve efficiency.
Gated MLA
A mechanism used in MoE models to dynamically route inputs to the most relevant experts, optimizing performance and resource usage.
FarSkip-Collective
A technique designed to reduce computational overhead in MoE models by skipping unnecessary computations during inference.
Sources · 1
Read next
More stories
TickrWireAI News Intelligence

We aggregate, verify, summarise and explain the latest artificial intelligence news from open, legal sources.

Daily AI digest

Top AI stories, summarised, in your inbox each morning.

© 2026 TickrWire. Summaries and analysis are AI-generated and may contain errors.