Qwen Releases Bilingual Open-Source Image Model
Alibaba's latest text-to-image generator, Qwen-Image 2512, is optimized for creating visuals from both English and Chinese prompts.
Alibaba's Qwen team has released a new text-to-image model, Qwen-Image 2512, expanding its suite of open-source AI tools. The generator is designed to create images from text prompts and is notable for its native support for both English and Chinese languages.
Released under the permissive Apache 2.0 license, Qwen-Image 2512 is available for both commercial and research applications without significant restrictions. This open approach continues Qwen's strategy of contributing foundational models to the community, following its well-regarded releases of large language models.
The model's bilingual capability is a key differentiator in a crowded field of image generators. By performing well with Chinese prompts, Qwen-Image addresses a significant need for tools that can create more culturally and linguistically specific visual content. This could empower developers and artists in Chinese-speaking regions to build more relevant applications. Interested users can find the model and usage details on its official Hugging Face repository.
Sources
- Visit
Qwen/Qwen-Image-2512
Hugging Face
More in Text → Image
SenseTime's SenseNova-U1.5 Unifies Vision Tasks
An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.
LLaDA-Image: A Fully Open 6B Image Generator
inclusionAI pairs a diffusion transformer with a frozen vision-language model and publishes the entire training recipe.
Kroma: An MIT-Licensed Text-to-Image Model for ComfyUI
Lodestones releases an open image generator built on the Krea 2 lineage, aimed squarely at ComfyUI workflows.
0 comments
No comments yet. Be the first to weigh in.