Meituan Releases Open-Source LongCat-Video Model
The Chinese tech giant has released a new MIT-licensed model capable of generating video from text, images, or by continuing existing clips.
Meituan, a major Chinese technology company, has entered the open-source video generation space with the release of LongCat-Video. The new model is available on Hugging Face with a permissive MIT license, allowing for broad use and modification.
LongCat-Video is a versatile diffusion model that creates short video clips from a variety of inputs. Its release adds another strong contender to the rapidly evolving field of open AI video, providing a flexible tool for developers and creators.
Core Capabilities
The model is designed to handle several distinct generation tasks, making it a multi-functional tool for video creation:
- Text-to-video: Generates a video sequence from a descriptive text prompt.
- Image-to-video: Animates a static source image based on a prompt.
- Video continuation: Extends an existing video clip by generating subsequent frames.
The model's permissive license and multifaceted approach make it a significant contribution to the open-source ecosystem. It provides researchers and commercial entities alike with a solid foundation for experimenting with and deploying generative video applications.
Sources
- Visit
meituan-longcat/LongCat-Video
Hugging Face
More in Text → Video

MiniMax Releases H3 Video Model on Hugging Face
The company's new diffusion model handles text-to-video and image-to-video, with support for joint audio-video generation.
LingBot-Video puts a 30B MoE behind embodied AI video
A DiT-based mixture-of-experts model activates just 3B parameters per step and ships under an Apache 2.0 license.

NVIDIA's Cosmos 3 Edge Brings World Models Closer
A new edge-optimized variant of NVIDIA's Cosmos world-model line aims to run generative video where the compute lives.
0 comments
No comments yet. Be the first to weigh in.