The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestGoogle DeepMind4 E4B
Google DeepMindAny-to-Any

Google Releases Gemma 4 E4B, a 4B Multimodal Model

The new 4-billion-parameter vision-language model brings image and text understanding to Google's popular open-source family.

Mar 2, 2026
NotableGemma
Gemma 4 E4B

Google DeepMind has expanded its open-source Gemma family with the release of Gemma 4 E4B, a new 4-billion-parameter multimodal model. This marks a significant step for the Gemma series, introducing vision capabilities to the previously text-focused lineup.

Unlike its predecessors, Gemma 4 E4B is a vision-language model (VLM) built to process and reason about images and text simultaneously. Following an "image-text-to-text" architecture, it can analyze visual information alongside textual prompts to generate relevant text-based responses. This allows it to handle tasks like visual question answering and image-based content generation.

The model's designation suggests a focus on efficiency at the 4-billion-parameter scale. By providing a relatively compact VLM, Google is targeting developers who need to build multimodal applications without relying on the extensive computational resources required by larger, proprietary models. This makes it a compelling option for use cases on consumer hardware or in other resource-constrained environments.

Gemma 4 E4B is available now on Hugging Face under the Gemma license, which allows for commercial use and distribution. Its release provides the open-source community with another powerful and accessible tool for building the next generation of AI applications.

Sources

  • google/gemma-4-E4B

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters4B
PrecisionBF16
ArchitectureGemma4ForConditionalGeneration
LicenseGEMMA
Downloads4.5M
Likes1.6K

Modalities

Any-to-AnyVision-LanguageText / LLM
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Any-to-Any

StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
SenseNova-U1.5-8B-MoT
SenseTime/Any-to-Any

SenseTime Releases SenseNova U1.5 8B Any-to-Any Model

The new 8B multimodal model handles text, images, and image editing within a single native architecture.

Aug 19, 2026