The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestMoonshot AIK2.6
Moonshot AIVision-Language

Kimi K2.6 tops closed models in coding test

Moonshot AI's open-weights mixture-of-experts model reportedly outperformed Claude, GPT-5.5, and Gemini on a programming challenge.

May 3, 2026
Major releaseOther

Moonshot AI has released Kimi K2.6, an open-weights mixture-of-experts model aimed squarely at software development, and early reports say it edged out some of the most capable proprietary systems on a programming challenge. According to a writeup at thinkpol.ca, the model outscored Claude, GPT-5.5, and Gemini in a head-to-head coding test.

The result matters less for any single benchmark number than for what it signals: an openly downloadable model competing with — and in this case reportedly beating — frontier closed offerings on code, one of the hardest and most commercially valuable tasks. If the outcome holds up across broader evaluations, it strengthens the case that open weights are no longer a step behind the leading labs.

Why it matters

  • Kimi K2.6 uses a mixture-of-experts design, which activates only part of the network per token to keep inference efficient at large scale.
  • Open weights mean developers can self-host, fine-tune, and audit the model rather than relying on an API.
  • A coding win against Claude, GPT-5.5, and Gemini puts pressure on the pricing and openness assumptions of closed providers.

A few caveats are worth keeping in mind. Single-challenge results can be noisy, and full details on parameter counts, context length, and licensing terms were not specified in the record. As always, independent replication across standard coding benchmarks will be the real test of whether K2.6's showing reflects durable capability or a favorable matchup.

Sources

  • Kimi K2.6 just beat Claude, GPT-5.5, and Gemini in a coding challenge

    Hacker News

    Visit

Get the model

Hacker News

Specs

LicenseOTHER
Downloads530K
Likes1.6K

Modalities

CodeText / LLMReasoning
4 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
inclusionAI/Vision-Language

LLaDA-UI Brings Diffusion Decoding to GUI Agents

inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.

Sep 8, 2026