The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestQwen · Alibaba3-VL
Qwen · AlibabaVision-Language

Alibaba Releases Qwen3-VL, an 8B Open-Source Vision Model

The latest vision-language model from the popular Qwen series is instruction-tuned and available under an Apache 2.0 license.

Oct 11, 2025
NotableApache 2.0
Qwen3-VL-8B-Instruct

Alibaba's Qwen team has launched Qwen3-VL-8B-Instruct, a new vision-language model (VLM) built on their latest Qwen3 architecture. This release adds powerful multimodal capabilities to the recently introduced Qwen3 family of open-source models.

As an instruction-tuned VLM, Qwen3-VL-8B is designed to understand and process both images and text simultaneously. It can perform a wide range of tasks that require visual reasoning, such as answering detailed questions about an image, generating captions, and identifying specific objects within a scene.

With 8 billion parameters, the model occupies a practical middle ground, offering strong performance without the demanding hardware requirements of much larger, proprietary systems. Its release provides developers and researchers with a capable tool for building multimodal applications, from enhanced chatbots to sophisticated content analysis systems.

The model is available under the permissive Apache 2.0 license, encouraging both academic and commercial use. Interested users can find the model weights and documentation on the Qwen Hugging Face repository.

Sources

  • Qwen/Qwen3-VL-8B-Instruct

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters8B
Size17.5 GB
PrecisionBF16
ArchitectureQwen3VLForConditionalGeneration
LicenseAPACHE-2.0
Downloads489.6K
Likes601

Modalities

Vision-Language
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
inclusionAI/Vision-Language

LLaDA-UI Brings Diffusion Decoding to GUI Agents

inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.

Sep 8, 2026