Hunyuan Image 3

HunyuanImage-3.0 multimodal text-to-image with top-tier quality and strong prompt adherence. Supports multiple sizes, seed, and JPEG/PNG output.

Cost: 6 credits

Input

Try:

The size of the generated image in pixels (width*height).

Set a specific seed for reproducible results; -1 uses a random seed.

The format of the output image.

Output

Generated content will appear here

Example Results

A teenage girl on a rooftop at sunset, ultra realistic details

Text-to-Image:

A teenage girl on a rooftop at sunset, ultra realistic details

A professional young woman walks down a street in New York, wearing a well-tailored grey wool blazer over a white silk camisole with matching wide-leg trousers. She is wearing delicate gold earrings and carrying a black leather tote bag. Her expression is confident, with a blurred city scene in the background.

Text-to-Image:

A professional young woman walks down a street in New York, wearing a well-tailored grey wool blazer over a white silk camisole with matching wide-leg trousers. She is wearing delicate gold earrings and carrying a black leather tote bag. Her expression is confident, with a blurred city scene in the background.

An old wizard with a long white beard, wearing a classic pointed wizard hat, sits comfortably in a wooden chair on top of his stone wizard tower. He leisurely smokes a pipe, with magical smoke swirling in the air around him, creating faint glowing shapes. The tower's top is high above the clouds, and a vast sky stretches out in the background, giving the scene a tranquil, isolated vibe. At the top of the image, large whimsical text reads: "I'm done procrastinating!" Further down, near the bottom-middle of the image, text in a similar whimsical style reads: "I'm straight up not working anymore!" The text is a key focal point, blending humorously into the laid-back, magical atmosphere. aidmaMJ6.1, hkmagic.

Text-to-Image:

An old wizard with a long white beard, wearing a classic pointed wizard hat, sits comfortably in a wooden chair on top of his stone wizard tower. He leisurely smokes a pipe, with magical smoke swirling in the air around him, creating faint glowing shapes. The tower's top is high above the clouds, and a vast sky stretches out in the background, giving the scene a tranquil, isolated vibe. At the top of the image, large whimsical text reads: "I'm done procrastinating!" Further down, near the bottom-middle of the image, text in a similar whimsical style reads: "I'm straight up not working anymore!" The text is a key focal point, blending humorously into the laid-back, magical atmosphere. aidmaMJ6.1, hkmagic.

Frequently Asked Questions

What is Hunyuan Image 3?
A state-of-the-art text-to-image model with strong multimodal understanding and premium image quality.
Which sizes are supported?
From 256×256 up to 1536×1536 depending on provider limits.
How do I get consistent results?
Use a fixed seed. The same prompt + seed will yield similar images.
What image format should I choose?
JPEG for smaller files, PNG for lossless quality and transparency workflows.
What makes it different from DiT-based models?
It uses a unified autoregressive architecture that jointly models text and image, improving contextual coherence, reasoning, and prompt adherence.
How large is the model?
It's a Mixture-of-Experts (MoE) with 64 experts and ~80B parameters in total, with ~13B parameters activated per token.
How strong is prompt adherence?
It demonstrates excellent semantic and compositional alignment in evaluations (e.g., SSAE and human GSB), delivering faithful results across diverse prompts.
Does it auto-expand sparse prompts?
Yes, the model leverages world knowledge to elaborate on minimal prompts. Tip: keep prompts concise and avoid contradictory instructions to steer expansion.
Any prompt writing tips?
Use subject + attributes + scene + lighting + camera/style cues. Add negative cues if needed. Avoid conflicting requirements and excessive style mixing.
Does it support image-to-image or multi-turn?
The open-source plan includes image-to-image and multi-turn interaction. This app currently exposes text-to-image parameters only.
What about license and usage?
HunyuanImage-3.0 is released under the tencent-hunyuan-community license. Review the license terms before commercial use.
How long does generation take?
Provider-side runtime varies with load and size (256–1536). Typical runs complete in seconds to tens of seconds; larger sizes can take longer.