The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestBaidu4.5-VL-Thinking
BaiduVision-Language

Baidu Releases Open Vision-Language MoE Model

The new ERNIE 4.5 VL model brings advanced multimodal reasoning to the open-source community with an efficient Mixture-of-Experts architecture.

Nov 7, 2025
NotableApache 2.0
ERNIE 4.5 VL 28B A3B Thinking

Chinese tech giant Baidu has released ERNIE 4.5 VL, a powerful new vision-language model available under a permissive open-source license. The model is designed for complex reasoning tasks that require understanding both images and text, positioning it as a capable new entry in the open multimodal space.

Efficient by Design

At its core, ERNIE 4.5 VL is a sparse Mixture-of-Experts (MoE) model. While it contains a total of 28 billion parameters, it only activates a fraction—around 3 billion—for any given inference task. This design, hinted at by the 'A3B' (Active 3 Billion) in its name, aims to provide the power of a much larger model while maintaining greater computational efficiency during use.

The model's full name, ERNIE 4.5 VL 28B A3B Thinking, emphasizes its focus on multi-step reasoning. It's built to analyze visual information and perform logical thinking, a challenging frontier for AI development.

By releasing this model under the Apache 2.0 license, Baidu is making a notable contribution to the open-source ecosystem. This gives researchers and developers a sophisticated, efficient, and freely usable foundation for building the next generation of multimodal applications.

Sources

  • baidu/ERNIE-4.5-VL-28B-A3B-Thinking

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters28B · MoE
Active params3B active
Size59.3 GB
PrecisionBF16
ArchitectureErnie4_5_VLMoeForConditionalGeneration
LicenseAPACHE-2.0
Downloads684
Likes543

Modalities

ReasoningVision-Language

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
inclusionAI/Vision-Language

LLaDA-UI Brings Diffusion Decoding to GUI Agents

inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.

Sep 8, 2026