Motif Releases 2B Open-Source Text-to-Video Model
The new Apache 2.0 licensed model uses a diffusion transformer architecture to offer a new open alternative for video generation research.

Motif Technologies has entered the open-source video generation space with the release of Motif-Video-2B, a new 2-billion-parameter model. Available under a permissive Apache 2.0 license, it provides a new, accessible tool for creating short video clips from both text and image prompts.
The model is built on a diffusion transformer architecture, an approach that has gained significant traction in generative AI for its strong performance in modeling complex data. Unlike many high-profile video systems that remain proprietary, Motif has made the model weights and code publicly available through its Hugging Face repository.
Why it matters
The release of Motif-Video-2B contributes to a growing but still-nascent ecosystem of open-source video generation models. Its relatively modest 2B parameter size makes it more approachable for researchers and developers who may not have access to the massive computational resources required by state-of-the-art commercial systems. By providing a transparent and modifiable foundation, the model can help accelerate experimentation and innovation in the broader community.
Sources
- Visit
Motif-Technologies/Motif-Video-2B
Hugging Face
More in Text → Video

MiniMax Releases H3 Video Model on Hugging Face
The company's new diffusion model handles text-to-video and image-to-video, with support for joint audio-video generation.
LingBot-Video puts a 30B MoE behind embodied AI video
A DiT-based mixture-of-experts model activates just 3B parameters per step and ships under an Apache 2.0 license.

NVIDIA's Cosmos 3 Edge Brings World Models Closer
A new edge-optimized variant of NVIDIA's Cosmos world-model line aims to run generative video where the compute lives.
0 comments
No comments yet. Be the first to weigh in.