The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestOpenBMB4.5
OpenBMBVision-Language

OpenBMB Releases Compact Multimodal Model MiniCPM-V 4.5

The new vision-language model from the open-source research group demonstrates strong OCR and video understanding capabilities in a small package.

Aug 24, 2025
NotableOther
MiniCPM-V-4.6

The open-source AI research group OpenBMB has released MiniCPM-V 4.5, a new and notably compact vision-language model (VLM). This model aims to deliver sophisticated multimodal understanding without requiring the massive computational resources often associated with leading-edge vision systems.

According to the release notes, the model demonstrates strong performance on tasks that have historically challenged even larger systems. Its key features include high-accuracy Optical Character Recognition (OCR), the ability to comprehend context across multiple images, and the capacity to understand video content—a significant step for a model in its size class.

Why it matters

The release of a smaller yet powerful VLM like MiniCPM-V is significant for developers working with limited hardware. Its efficiency opens up possibilities for on-device applications and more accessible multimodal AI research, lowering the barrier to entry for building sophisticated vision-based tools.

The model is now available for download and experimentation. Interested developers can find all the resources on the Hugging Face Hub. The model is available under a custom license, so users should review the terms before deployment in production environments.

Sources

  • openbmb/MiniCPM-V-4_5

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size17.4 GB
PrecisionBF16
ArchitectureMiniCPMV
LicenseOTHER
Downloads456.2K
Likes1.2K

Modalities

Vision-Language
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
inclusionAI/Vision-Language

LLaDA-UI Brings Diffusion Decoding to GUI Agents

inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.

Sep 8, 2026