The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestWeiboAI3B
WeiboAIReasoning

Weibo AI Releases VibeThinker-3B, a Compact Reasoning Model

The new 3-billion-parameter model from the Chinese tech giant focuses on challenging benchmarks in mathematics, coding, and graduate-level questions.

Jun 12, 2026
NotableOther
VibeThinker-3B

Chinese technology company Weibo has introduced VibeThinker-3B, a new small language model focused on advanced reasoning. At just three billion parameters, the model is part of a growing class of highly efficient models designed to deliver specialized performance without the massive computational overhead of their larger counterparts.

According to the release card on Hugging Face, VibeThinker-3B was developed to excel at specific, difficult tasks. The creators highlight its performance on benchmarks that test mathematical ability (GSM8K, MATH), code generation (HumanEval), and graduate-level, Google-proof question answering (GPQA), indicating a focus on deep, domain-specific problem-solving rather than general conversation.

A Niche Specialist

The model's deliberate focus on reasoning-intensive domains is what sets it apart. While many small models aim for broad competence, VibeThinker-3B is positioned as a specialist. This strategy allows smaller models to potentially outperform much larger ones on targeted tasks, making them valuable components for applications requiring reliable logic, math, or code intelligence.

Why it matters: The release of specialized models like VibeThinker-3B demonstrates a maturing ecosystem where developers can choose the right tool for the job. Instead of relying on a single, monolithic model, teams can deploy smaller, more cost-effective models fine-tuned for specific needs. Prospective users should note the model is released under a custom license, and its terms should be reviewed before use.

Sources

  • WeiboAI/VibeThinker-3B

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters3B
Size6.2 GB
PrecisionBF16
ArchitectureQwen2ForCausalLM
LicenseOTHER
Downloads108.5K
Likes813

Modalities

ReasoningText / LLMCode

0 comments

No comments yet. Be the first to weigh in.

More in Reasoning

DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents

The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.

Jul 31, 2026
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Refreshes V4-Flash With New 0731 Checkpoint

The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.

Jul 31, 2026
K-EXAONE 2.0 750B-A37B
LGAI EXAONE/Text / LLM

LG AI Research debuts K-EXAONE 2.0, a 750B MoE model

The new mixture-of-experts model activates 37B parameters per token and targets English, Korean, and Spanish reasoning tasks.

Jul 29, 2026