HunyuanImage

An open-source, commercial-grade text-to-image AI model delivering photorealistic images with advanced semantic understanding and world knowledge reasoning.

Overview

Hunyuan Image 3.0 is a state-of-the-art text-to-image generation model developed and open-sourced by Tencent in September 2025. It employs a revolutionary unified autoregressive multimodal architecture that deeply integrates text and image modalities, enabling superior semantic understanding and photorealistic image synthesis. The model features a Mixture of Experts (MoE) design with 64 experts and a total of 80 billion parameters, of which 13 billion are activated per token during inference. This architecture supports complex semantic comprehension, bilingual input (Chinese and English), and advanced world knowledge reasoning, allowing it to generate detailed, contextually rich images from sparse or complex prompts. It also incorporates advanced technologies such as prompt enhancement, refiner models, and hierarchical semantic processing to optimize image quality and text-image alignment.