LightOn Releases OCR-2, a 1B Document AI Model
The new vision model from the Paris-based AI lab uses Mistral architecture to extract text and structure from complex documents like PDFs and forms.

Parisian AI company LightOn has released LightOnOCR-2, a new 1-billion-parameter vision language model specialized in document understanding. The model is designed to perform optical character recognition (OCR) on complex documents, extracting not just text but also structural information.
Unlike simple OCR tools, LightOnOCR-2 is built to parse challenging layouts like tables, forms, and multi-column PDFs. This makes it suitable for enterprise automation tasks such as processing invoices or digitizing records, a domain often dominated by proprietary, API-gated services.
A Mistral-based Architecture
The model features a Transformer-based encoder-decoder architecture. In an interesting design choice, its decoder was initialized using a subset of weights from Mistral-7B-v0.1, allowing it to leverage the powerful language capabilities of the popular open model while maintaining a much smaller, more efficient footprint.
The complete model weights and code are available on the Hugging Face Hub for developers to download and use. It's released under a custom LightOnAI-OpenRAIL-M license, which permits commercial use but includes some use-case restrictions common to Responsible AI licenses.
Sources
- Visit
lightonai/LightOnOCR-2-1B
Hugging Face
More in Vision-Language
Agnes-3.0-Flash arrives as a multimodal reasoning model
The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.
SenseTime's SenseNova-U1.5 Unifies Vision Tasks
An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.
LLaDA-UI Brings Diffusion Decoding to GUI Agents
inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.
0 comments
No comments yet. Be the first to weigh in.