ImageBench

vs

192 evaluations across 6 categories

Benchmark V1 verdicts are produced by VLM judges and can contain mistakes. Treat PASS/FAIL labels as machine-assisted assessments, and inspect the images yourself. Learn more about the methodology.

Generation Details

Source-backed model context, size, cost, and request settings for this ImageBench V1 run.

openai/gpt-image-2

API

GPT Image 2 is OpenAI's current image generation and editing model for fast, high-quality image creation, with flexible image sizes and high-fidelity image inputs through the OpenAI Images API.

Maker
OpenAI
Family
GPT Image
Model Size
not disclosed
unknown
Cost
token-based
high
Run Target
openai/gpt-image-2
Effective Request
model: gpt-image-2 · size: 1024x1024 · quality: auto

fal/qwen-image-2-pro

API

Qwen Image 2.0 Pro is Alibaba's Qwen (Tongyi Lab) flagship hosted text-to-image model, a top-10 model on the Artificial Analysis image arena known for best-in-class typography, infographics, and poster/layout rendering. Distinct from the older local Qwen-Image checkpoint on this board; released ~April 2026.

Maker
Alibaba (Qwen)
Family
Qwen Image 2.0
Model Size
not disclosed
unknown
Cost
$0.075 / image
high
Run Target
fal/fal-ai/qwen-image-2/pro/text-to-image
Effective Request
image_size: square_hd · seed: 7
80.0vs66.9
Overall
87%vs71%
Capability
73.0vs63.1
Est. Preference
45.3svs12.8s
Avg Latency
Text Rendering100%100%Spatial Reasoning91%75%Human realism69%38%Truthfulness74%59%Professional Studio100%93%Graphical design100%88%Preference73%63%Latency0%25%

All 192 generations

Text Rendering100%vs100%

openai/gpt-image-2fal/qwen-image-2.0-pro

Typography Style100%vs100%

Writing accuracy100%vs100%

Spatial Reasoning91%vs75%

openai/gpt-image-2fal/qwen-image-2.0-pro

Attributes Binding100%vs89%

Compositionality100%vs89%

Counting89%vs67%

Negation100%vs89%

Relative Position83%vs83%

Scale & Proportions78%vs33%

Human realism69%vs38%

openai/gpt-image-2fal/qwen-image-2.0-pro

Faces & Expressions92%vs92%

Full Body17%vs17%

Hands83%vs0%

Multi-Subject100%vs50%

Truthfulness74%vs59%

openai/gpt-image-2fal/qwen-image-2.0-pro

Photorealism100%vs33%

Physics & Reflections58%vs50%

World Knowledge83%vs75%

Professional Studio100%vs93%

openai/gpt-image-2fal/qwen-image-2.0-pro

Camera & Lighting100%vs83%

Color Precision100%vs100%

Photorealism100%vs100%

Graphical design100%vs88%

openai/gpt-image-2fal/qwen-image-2.0-pro

Layout & Design100%vs78%

Data Visualisation100%vs67%

Style Diversity100%vs100%