Meta's Muse Glimmer 30B Targets Local Agentic Coding
A 30-billion-parameter multimodal model built to run locally, released under Apache 2.0 with an eye on agentic coding workflows.
The company's newest flagship targets reasoning and coding while keeping a permissive open-source license.
A 30-billion-parameter multimodal model built to run locally, released under Apache 2.0 with an eye on agentic coding workflows.
Alibaba's Qwen team pushes its largest sparse model yet, activating 95 billion parameters per token from a 2.4-trillion-parameter pool.
Alibaba's latest Qwen release pairs image understanding with text under a permissive Apache-2.0 license.
inclusionAI's new text model uses a mixture-of-experts design to keep compute low while shipping under an Apache-2.0 license.
The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.
The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.
A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.
The new open model separates musical structure from sound, generating long-form tracks from text prompts with an explicit planning stage.
A single discrete autoregressive model handles speech, voice design, sound effects, and music generation.
Momentum
Zhipu AI
Text / LLM · 2.4M downloads
DeepSeek
Text / LLM · 2.2M downloads
Google DeepMind
Any-to-Any · 3.6M downloads
rednote-hilab
Vision-Language · 710.9K downloads
DeepSeek
Vision-Language · 668.3K downloads
Benchmarks
| # | Model | Avg. |
|---|---|---|
| 1 | MaziyarPanahi/calme-3.2-instruct-78b | 52.1 |
| 2 | MaziyarPanahi/calme-3.1-instruct-78b | 51.3 |
| 3 | dfurman/CalmeRys-78B-Orpo-v0.1 | 51.2 |
| 4 | MaziyarPanahi/calme-2.4-rys-78b | 50.8 |
| 5 | huihui-ai/Qwen2.5-72B-Instruct-abliterated | 48.1 |
| 6 | Qwen/Qwen2.5-72B-Instruct Qwen · Alibaba | 48.0 |
An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.
A 122B mixture-of-experts model trained with reinforcement learning claims state-of-the-art results on Terminal-Bench.

The company's new diffusion model handles text-to-video and image-to-video, with support for joint audio-video generation.
The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.
Alibaba's next Qwen release pairs a large parameter pool with a tiny active footprint, promising speed without the full compute bill.
The new mixture-of-experts model is billed as more capable than V4 Pro while costing less to run, and ships under an MIT license.
The Chinese AI lab says its stealth-tested system belongs to the GLM series and will ship with open weights under an MIT license.