Tencent's T1 Targets Long-Horizon Terminal Work
A 122B mixture-of-experts model trained with reinforcement learning claims state-of-the-art results on Terminal-Bench.
Company
Releases
A 122B mixture-of-experts model trained with reinforcement learning claims state-of-the-art results on Terminal-Bench.
The company's next-generation Hunyuan language model arrives as an early preview with a permissive license and a mixture-of-experts design.
The WeChat team releases a 9-billion-parameter multimodal embedding model that maps three modalities into one shared vector space.
The new open-weights model handles zero-shot TTS alongside enhancement and separation, aiming to be a broad speech toolkit rather than a single-purpose voice engine.
A 27B vision-language model built to see and operate graphical interfaces, tested on OSWorld and WindowsAgentArena.
The company's latest mixture-of-experts model arrives as an openly licensed conversational LLM on Hugging Face.
A lightweight image-editing framework claims results rivaling 10B-scale models, and it's already running in the browser.
The 1.8 billion-parameter model from the Chinese tech giant is designed for high-quality translation across a wide range of language pairs.
The new HY-Embodied 0.5 is a vision-language model designed specifically for multi-object tracking in dynamic, real-world environments.
Built on their HunyuanVideo-1.5 architecture, the new model synthesizes video by combining multiple static images and text prompts into a cohesive narrative.
The new model from Tencent's Hunyuan team generates dynamic video and reconstructs 3D environments using a single static picture.
The new diffusion model generates short video clips from text and image prompts, adding another major player to the open video space.
The new vision-language model from Tencent Hunyuan offers a compact, end-to-end solution for optical character recognition.
The new text-to-image generator from the Chinese tech giant uses a Mixture-of-Experts architecture for more efficient and detailed image creation.
The new text-to-image generator from the Chinese tech giant uses a Mixture-of-Experts architecture for improved efficiency and output quality.
The new text-to-image model uses a novel rejection sampling technique to align Stable Diffusion XL more closely with human aesthetic preferences.
The new text-to-image model from the Chinese tech giant is designed to understand both Chinese and English prompts at high resolutions.
The new model from Tencent AI Lab generates temporally and spatially consistent video sequences from a single image, enabling virtual exploration of static scenes.
The new Hunyuan-GameCraft 1.0 is an open image-to-video model that generates interactive game-like scenes with precise camera control.
The new Apache 2.0-licensed generator uses a Mixture-of-Experts architecture and is available in the popular Diffusers library format for easier integration.