The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestBaiduv1
BaiduVision-Language

Baidu releases Unlimited-OCR under permissive MIT license

The Chinese tech giant's multilingual vision-language model targets text extraction across languages and document types.

Jun 19, 2026
NotableMIT
Unlimited-OCR

Baidu has published Unlimited-OCR on Hugging Face, a multilingual vision-language model aimed at optical character recognition. The release lands under the MIT license, one of the most permissive options available, meaning developers can use, modify, and ship the model commercially with minimal restrictions.

Unlimited-OCR is built as a vision-language system rather than a traditional OCR pipeline. That approach has become increasingly common for document understanding, where models read images and produce structured or plain text while drawing on broader language reasoning to handle layout, context, and multiple scripts.

Why it matters

OCR remains one of the most practical and widely deployed AI tasks, underpinning everything from document digitization to data entry and accessibility tools. A few points stand out about this release:

  • The MIT license lowers the barrier for commercial adoption and downstream fine-tuning.
  • A multilingual focus suggests the model is intended to work across scripts rather than English-only text.
  • Backing from Baidu, a major player in Chinese-language AI, adds weight to the multilingual claim.

Baidu has not yet detailed parameter counts, context length, or benchmark results in the available record, so independent evaluation will be the real test of how Unlimited-OCR compares to established open OCR systems. For now, the model is available to download and try directly from its Hugging Face repository.

Sources

  • baidu/Unlimited-OCR

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size6.7 GB
PrecisionBF16
ArchitectureUnlimitedOCRForCausalLM
LicenseMIT
Downloads2.6M
Likes3.8K

Modalities

Vision-Language

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Inkling Small
Thinkingmachines/Vision-Language

Thinking Machines Debuts Inkling Small, a Compact Multimodal MoE

The Apache-2.0 model brings mixture-of-experts efficiency to image, audio, and text tasks in a smaller footprint.

Jul 27, 2026
Mage-VL
Microsoft/Vision-Language

Microsoft's Mage-VL Streams Video Natively

A codec-native multimodal foundation model aims to understand live video and vision-language input in real time.

Jul 26, 2026
Apertus v1.5 70B
Swiss Ai/Text / LLM

Apertus v1.5 70B arrives with an Apache-2.0 license

Switzerland's open-model effort ships a 70-billion-parameter, multilingual and multimodal system that anyone can use, modify, and deploy.

Jul 24, 2026