The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestGoogle DeepMind4
Google DeepMindAny-to-Any

Google's Gemma 4 Arrives with Any-to-Any Multimodal Skills

The new 2-billion-parameter model from DeepMind can process text, vision, and audio, making it a versatile and efficient foundation for developers.

Mar 2, 2026
Major releaseGemma
Gemma 4 E2B

Google DeepMind has released Gemma 4 E2B IT, the first model in a new generation of its open-weights AI family. This compact, 2-billion-parameter model is distinguished by its native "any-to-any" multimodality, capable of processing and generating combinations of text, vision, and audio.

The "E2B" in the name likely stands for "Efficient 2 Billion," highlighting the model's focus on performance within a small footprint. As an instruction-tuned (-it) variant, it's optimized for chat and direct-response tasks and features a context window of 8192 tokens.

Why It Matters

Gemma 4's release marks a significant step for developers seeking powerful AI that can run efficiently on local hardware or in resource-constrained cloud environments. By integrating text, vision, and audio capabilities into a single, compact model, Google is making sophisticated, multi-sensory AI more accessible and lowering the barrier for building complex applications.

The model is available now for researchers and developers on Hugging Face. It is released under the Gemma license terms, which permit commercial use but include specific restrictions, continuing Google's strategy of providing capable but controlled open-source tools.

Sources

  • google/gemma-4-E2B-it

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters2B
PrecisionBF16
ArchitectureGemma4ForConditionalGeneration
LicenseGEMMA
Downloads3.6M
Likes964

Modalities

Any-to-AnyVision-LanguageText / LLM
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Any-to-Any

StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
SenseNova-U1.5-8B-MoT
SenseTime/Any-to-Any

SenseTime Releases SenseNova U1.5 8B Any-to-Any Model

The new 8B multimodal model handles text, images, and image editing within a single native architecture.

Aug 19, 2026