tencent/HunyuanImage-3.0
Text-to-Image β’ 83B β’ Updated β’ 3.49k β’ β’ 1.13k
Generate high-quality speech from text with optional voice cloning
OmniParser, turn your LLM into GUI agent
Generate depth video from input video
Audio Conditioned LipSync with Latent Diffusion Models