ImageBench

ImageBench V1 —

192 evaluations across 6 categories

Benchmark V1 verdicts are produced by VLM judges and can contain mistakes. Treat PASS/FAIL labels as machine-assisted assessments, and inspect the images yourself. Learn more about the methodology.

Generation Details

Source-backed model context, size, cost, and request settings for this ImageBench V1 run.

fal/hunyuan-image-3

API

HunyuanImage 3.0 is Tencent's flagship open-weight text-to-image model — an 80B-parameter Mixture-of-Experts (~13B active), among the largest open image models. It emphasizes strong prompt following and legible in-image text, and reaches roughly ELO ~1125 on the Artificial Analysis image arena. Brings a new maker (Tencent) to the benchmark.

Maker
Tencent
Family
HunyuanImage 3.0
Model Size
80B MoE (~13B active)
high
Cost
~$0.09 / image
medium
Run Target
fal/fal-ai/hunyuan-image/v3/text-to-image
Effective Request
image_size: square_hd · seed: 7 · num_inference_steps: 28 · guidance_scale: 7.5
57.7
Overall
52%
Capability
63.2
Est. Preference
100
Pass
92
Fail
24.4s
Avg Latency
14.0s
Min Latency
159.0s
Max Latency
Text Rendering53%Spatial Reasoning46%Human realism41%Truthfulness44%Professional Studio85%Graphical design58%Preference63%Latency6%

All 192 generations

Text Rendering53%

Typography Style100%

Writing accuracy42%

Spatial Reasoning46%

Attributes Binding56%

Compositionality67%

Counting22%

Negation22%

Relative Position58%

Scale & Proportions44%

Human realism40%

Faces & Expressions58%

Full Body0%

Hands75%

Multi-Subject17%

Truthfulness44%

Photorealism67%

Physics & Reflections50%

World Knowledge33%

Professional Studio85%

Camera & Lighting92%

Color Precision92%

Photorealism33%

Graphical design58%

Data Visualisation0%

Layout & Design44%

Style Diversity83%