Datalab To/Vision-Language
LIFT: A Qwen3.5-Based VLM for PDF-to-JSON Extraction
Datalab's new open vision-language model targets structured data extraction from documents, turning messy PDFs into clean JSON.
Company
Releases
Datalab's new open vision-language model targets structured data extraction from documents, turning messy PDFs into clean JSON.
The new vision-language model from Datalab is fine-tuned from Qwen2-VL to specialize in extracting text and structure from complex documents.