The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestQwen · Alibaba3-VL
Qwen · AlibabaVision-Language

Qwen Releases 30B MoE Vision Model, Qwen3-VL

The new open-source model from Alibaba uses a Mixture-of-Experts architecture to make its powerful vision-language capabilities more efficient to run.

Sep 30, 2025
Major releaseApache 2.0
Qwen3-VL-8B-Instruct

Alibaba's Qwen team has released Qwen3-VL, a new open-source vision-language model (VLM) that combines high performance with computational efficiency. This instruction-tuned model is designed to understand and process both text and images, making it suitable for a wide range of multimodal tasks.

The model's key innovation is its Mixture-of-Experts (MoE) architecture. While it contains a total of 30 billion parameters, only 3 billion are activated during inference for any given input. This design allows it to achieve the performance associated with a much larger model while maintaining the speed and lower resource requirements of a smaller one, a significant advantage for developers and researchers.

As an instruction-tuned model, Qwen3-VL is optimized for conversational and task-oriented applications. It can follow complex commands that involve analyzing visual content, such as answering detailed questions about an image or generating descriptive captions. This makes it a powerful tool for building more sophisticated AI assistants and applications.

The model is released under the permissive Apache 2.0 license, encouraging broad adoption for both academic and commercial projects. Full details and model weights are available on its Hugging Face repository.

Sources

  • Qwen/Qwen3-VL-30B-A3B-Instruct

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters30B · MoE
Size62.1 GB
PrecisionBF16
ArchitectureQwen3VLMoeForConditionalGeneration
LicenseAPACHE-2.0
Downloads489.6K
Likes601

Modalities

Any-to-AnyVision-Language
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
inclusionAI/Vision-Language

LLaDA-UI Brings Diffusion Decoding to GUI Agents

inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.

Sep 8, 2026