The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestGoogle DeepMind4
Google DeepMindAny-to-Any

Google DeepMind's Gemma 4 Goes Multimodal and MoE

The new open-weights family adds a mixture-of-experts design, encoder-free multimodal inputs, and an optional thinking mode.

Jul 1, 2026
Major releaseGemma

Google DeepMind has introduced Gemma 4, the latest generation of its open-weights model family, according to the Gemma 4 Technical Report. The release marks a notable architectural shift, moving the family toward a mixture-of-experts (MoE) design while broadening its native support for text, images, and reasoning-heavy tasks.

The headline changes are structural. Gemma 4 adopts an MoE approach, which activates only a subset of parameters per token to improve efficiency relative to dense models of comparable capacity. The family is also described as encoder-free for multimodal inputs, suggesting a more unified path for handling images and text without a separate vision encoder stage.

What's new in Gemma 4

  • A mixture-of-experts architecture across the family
  • Encoder-free multimodal handling for text and vision
  • An optional thinking mode for step-by-step reasoning
  • Support for long context

The thinking mode places Gemma 4 alongside a growing set of models that expose explicit reasoning behavior, letting developers trade extra compute for stronger performance on complex problems. Combined with long-context support, that positions the family for tasks that span large documents or multi-step workflows.

Gemma 4 continues to ship under the Gemma license, keeping the weights accessible for developers and researchers who want to build on or fine-tune the models directly. Why it matters: an open MoE multimodal family with reasoning controls narrows the gap between open weights and the capabilities typically reserved for closed frontier systems, giving teams more room to self-host and customize.

Sources

  • Gemma 4 Technical Report

    HF Papers

    Visit

Get the model

HF Papers

Specs

LicenseGEMMA
Downloads3M
Likes1.4K

Modalities

Any-to-AnyText / LLMVision-LanguageReasoning
6 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Any-to-Any

Inkling Small
Thinkingmachines/Vision-Language

Thinking Machines Debuts Inkling Small, a Compact Multimodal MoE

The Apache-2.0 model brings mixture-of-experts efficiency to image, audio, and text tasks in a smaller footprint.

Jul 27, 2026
A.X-K2 Raon Speech 21B-A3B
KRAFTON/Any-to-Any

KRAFTON releases A.X-K2 Raon speech MoE model

The game maker's new open model blends text-to-speech and speech recognition in a single 21B mixture-of-experts system with just 3B active parameters.

Jul 27, 2026
Mage-VL
Microsoft/Vision-Language

Microsoft's Mage-VL Streams Video Natively

A codec-native multimodal foundation model aims to understand live video and vision-language input in real time.

Jul 26, 2026