The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestAmd1.0
AmdReasoning

AMD's Instella-MoE Brings Reasoning to ROCm Hardware

A new open mixture-of-experts model with 16B total parameters and just 3B active is tuned to run on AMD's own accelerator stack.

Jul 23, 2026
NotableOther
Instella-MoE 16B A3B Think

AMD has released Instella-MoE 16B A3B Think, an open-weight reasoning model built as a mixture of experts. The design pairs 16 billion total parameters with roughly 3 billion active at inference, a structure meant to keep compute costs modest while retaining the capacity of a larger network.

The "Think" label signals the model's focus: it is positioned as a reasoning-oriented system rather than a general chat assistant, joining a growing field of open models that expose step-by-step problem solving. It handles text generation alongside its reasoning workload.

Why it matters

The most notable angle here is the hardware. AMD explicitly optimizes Instella-MoE for its ROCm software stack, giving developers on AMD accelerators a first-party option rather than defaulting to models tuned for competing platforms.

  • 16B total parameters, ~3B active per token
  • Reasoning and text generation focus
  • Optimized for AMD's ROCm ecosystem
  • Open weights available on Hugging Face

As an initial 1.0 release, Instella-MoE reflects AMD's continued push to build a credible open-model presence tied to its own silicon. The weights and details are published on Hugging Face under a custom license, so teams weighing it for production should review the terms before deploying.

Sources

  • amd/Instella-MoE-16B-A3B-Think

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters16B · MoE
Active params3B active
Size31.7 GB
PrecisionBF16
ArchitectureInstellaMoEForCausalLM
LicenseOTHER
Downloads2.1K
Likes143

Modalities

Text / LLMReasoning

0 comments

No comments yet. Be the first to weigh in.

More in Reasoning

DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents

The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.

Jul 31, 2026
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Refreshes V4-Flash With New 0731 Checkpoint

The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.

Jul 31, 2026
K-EXAONE 2.0 750B-A37B
LGAI EXAONE/Text / LLM

LG AI Research debuts K-EXAONE 2.0, a 750B MoE model

The new mixture-of-experts model activates 37B parameters per token and targets English, Korean, and Spanish reasoning tasks.

Jul 29, 2026