---
title: "GPT Image 2 Benchmark Results"
description: "OpenAI's GPT Image 2: the successor to GPT Image 1.5, run here at the high quality tier."
url: "https://img.ly/ai-benchmarks/models/gpt-image-2/"
type: "benchmark-model"
provider: "OpenAI"
tier: "extended"
costPerImageUsd: 0.1551
---

> This is the markdown version of [GPT Image 2 Benchmark Results](https://img.ly/ai-benchmarks/models/gpt-image-2/). For all pages in one file, see [llms-full.txt](https://img.ly/llms-full.txt). For an index of all available pages, see [llms.txt](https://img.ly/llms.txt).

---

# GPT Image 2

OpenAI's GPT Image 2: the successor to GPT Image 1.5, run here at the high quality tier.

GPT Image 2 is OpenAI's successor to GPT Image 1.5, new to this edition, and it is the sharpest illustration in this dataset of a newer model being a worse fit for print. It takes the best brand-color score in the field at 4.11 of 5 and sits second on prompt adherence at 4.88, with a clean 5.00 on text. What it lost is the alpha channel: its predecessor scores 4.22 on transparency, GPT Image 2 scores zero and paints a checkerboard into the pixels instead. It is also the most expensive model measured, at $0.155 per image against $0.140 for GPT Image 1.5, and the slowest of the OpenAI line at 110 seconds. Route brand-critical work here and keep it away from anything that needs a cut-out.

## Why it stands out

- Best brand-color accuracy in the field: 4.11 of 5 on CIEDE2000 distance to the requested hex.
- Second on prompt adherence at 4.88 and a clean 5.00 on text.
- Lost the alpha channel its predecessor had: 0.00 against 4.22 for GPT Image 1.5.
- The most expensive model in the benchmark at $0.155 per image, and 110 seconds per generation.

- Provider: OpenAI
- Tier: extended
- Benchmarked endpoint: fal:fal-ai/gpt-image-2
- Available on the IMG.LY AI Gateway: no
- Cost per image: $0.155
- p50 latency: 110.1s
- Images benchmarked: 111
- Licensing: Commercial use via OpenAI API terms (served through fal).

## Scores by criterion (auto tier)

- Transparency: 0.0 / 5 (9 scores across 9 runs)
- Color Accuracy: 4.1 / 5 (9 scores across 9 runs)
- Composition: 4.2 / 5 (9 scores across 9 runs)
- Cost: 3.0 / 5 (111 scores across 111 runs)
- Latency: 1.2 / 5 (111 scores across 111 runs)
- Prompt Adherence: 4.9 / 5 (111 scores across 111 runs)
- Resolution: 2.0 / 5 (111 scores across 111 runs)
- Text Accuracy: 5.0 / 5 (24 scores across 24 runs)

Blended overall (each scored criterion one equal vote): 3.04 / 5.
Includes the interim Claude-VLM quality tier (prompt adherence, composition);
not yet blind expert-panel reviewed. Because the blend is unweighted, cost and
latency pull cheap, fast models above premium flagships. The use-case pages
reweight the same measurements by job.

## Scores by category

- Typography & Text Rendering: 3.2 / 5
- Logos, Icons & Vector-Style: 2.8 / 5
- Brand-Color Fidelity: 3.1 / 5
- Transparency & Cutouts: 2.2 / 5
- Composition & Negative Space: 3.0 / 5
- Spatial & Compositional Adherence: 2.8 / 5
- Human Subjects: Faces & Hands: 2.7 / 5
- Product & E-Commerce Staging: 2.8 / 5
- Consistency & Repeatability: 2.6 / 5
- Print Detail & Resolution: 2.8 / 5
- Style Adherence & Art Direction: 2.8 / 5
- Print-Production Graphics: 3.2 / 5
- UI & Design Mockups: 3.2 / 5
- Diagrams & Data Viz: 3.2 / 5
- Structured Spec Adherence: 2.9 / 5

## Benchmarks featuring this model

**Typography & Text Rendering**

- [Wordmark](https://img.ly/ai-benchmarks/prompts/t04-wordmark.md)
- [Poster Headline](https://img.ly/ai-benchmarks/prompts/t01-poster-headline.md)
- [Product Label](https://img.ly/ai-benchmarks/prompts/t02-dense-label.md)
- [CJK Neon Sign](https://img.ly/ai-benchmarks/prompts/t05-cjk-sign.md)

**Logos, Icons & Vector-Style**

- [Mascot Logo](https://img.ly/ai-benchmarks/prompts/l02-mascot.md)
- [Icon Set](https://img.ly/ai-benchmarks/prompts/l01-icon-set.md)
- [Monochrome Emblem](https://img.ly/ai-benchmarks/prompts/l03-emblem.md)

**Brand-Color Fidelity**

- [Two-Color Illustration](https://img.ly/ai-benchmarks/prompts/b01-two-color.md)
- [Duotone Portrait](https://img.ly/ai-benchmarks/prompts/b03-duotone.md)
- [ColorChecker Chart](https://img.ly/ai-benchmarks/prompts/b04-colorchecker.md)
- [Smooth Gradient](https://img.ly/ai-benchmarks/prompts/b05-gradient-banding.md)

**Transparency & Cutouts**

- [Die-Cut Sticker](https://img.ly/ai-benchmarks/prompts/c01-sticker.md)
- [T-Shirt Graphic](https://img.ly/ai-benchmarks/prompts/c02-shirt-graphic.md)
- [Isolated Cutout](https://img.ly/ai-benchmarks/prompts/c03-isolated-object.md)

**Composition & Negative Space**

- [Banner with Negative Space](https://img.ly/ai-benchmarks/prompts/n01-banner.md)
- [Vertical Story Background](https://img.ly/ai-benchmarks/prompts/n02-story-bg.md)
- [Centered Product Margin](https://img.ly/ai-benchmarks/prompts/n03-margin-product.md)

**Spatial & Compositional Adherence**

- [Object Counting & Placement](https://img.ly/ai-benchmarks/prompts/x01-objects.md)
- [Scene Relations](https://img.ly/ai-benchmarks/prompts/x02-scene-relations.md)
- [Negation: Clean Background](https://img.ly/ai-benchmarks/prompts/x03-negation.md)

**Human Subjects: Faces & Hands**

- [Hand Holding a Sphere](https://img.ly/ai-benchmarks/prompts/h03-sphere-grip.md)

**Product & E-Commerce Staging**

- [Cosmetics Hero](https://img.ly/ai-benchmarks/prompts/p01-cosmetics.md)
- [Floating Sneaker](https://img.ly/ai-benchmarks/prompts/p02-sneaker.md)

**Consistency & Repeatability**

- [Character Series: Greenhouse](https://img.ly/ai-benchmarks/prompts/s01-character-greenhouse.md)
- [Character Series: Desk](https://img.ly/ai-benchmarks/prompts/s02-character-desk.md)
- [Character Series: Market](https://img.ly/ai-benchmarks/prompts/s03-character-market.md)

**Print Detail & Resolution**

- [Fine-Line Mandala](https://img.ly/ai-benchmarks/prompts/r01-mandala.md)
- [Illustrated Map](https://img.ly/ai-benchmarks/prompts/r02-map.md)

**Style Adherence & Art Direction**

- [Mid-Century Travel Poster](https://img.ly/ai-benchmarks/prompts/a01-travel-poster.md)
- [Isometric Office](https://img.ly/ai-benchmarks/prompts/a03-isometric.md)

**Print-Production Graphics**

- [Greeting Card](https://img.ly/ai-benchmarks/prompts/g01-greeting-card.md)
- [Seamless Pattern Tile](https://img.ly/ai-benchmarks/prompts/g03-pattern-tile.md)
- [Stationery Mockup](https://img.ly/ai-benchmarks/prompts/g04-stationery-mockup.md)
- [Three-Panel Comic](https://img.ly/ai-benchmarks/prompts/g05-comic-strip.md)

**UI & Design Mockups**

- [E-Commerce Product Page](https://img.ly/ai-benchmarks/prompts/ui03-pdp.md)

**Diagrams & Data Viz**

- [Labeled Bar Chart](https://img.ly/ai-benchmarks/prompts/d01-bar-chart.md)

**Structured Spec Adherence**

- [Structured JSON Scene](https://img.ly/ai-benchmarks/prompts/j01-json-scene.md)


## Frequently asked questions

**Why am I seeing three samples per benchmark?**

Image generation is stochastic: the same prompt produces different images on every run, so a single sample measures luck, not ability. Every model runs each benchmark three times with fixed seeds (1111, 2222, 3333), or three unseeded samples where the API accepts no seed. Scores average all three samples, and the consistency benchmarks measure the variation itself.

**Is GPT Image 2 good at diagrams & data viz?**

Diagrams & Data Viz is where GPT Image 2 scores best against its own other categories, which is not the same as leading the field there. The per-category breakdown on this page shows where it is strong and weak, so use the category scores, not the blended overall, to judge it for a specific job.

**How much does GPT Image 2 cost per image?**

GPT Image 2 costs $0.155 per image at the endpoint and default parameters we benchmarked, shown in the stats above. Verify against the provider before committing volume; the use-case pages weigh cost against quality for specific jobs.

**Can I use GPT Image 2 output in production without editing?**

Not reliably. Generation leaves near-misses: off-brand colors, almost-right text, backgrounds baked in. The IMG.LY AI Editor lets your users refine GPT Image 2's output to production quality with background removal, brand kits and editable text on a real canvas.

**How does GPT Image 2 compare to other models?**

The model rankings place GPT Image 2 against every other model on the same prompts, and the comparison pages put it head to head with a specific rival. Because the blended score is unweighted, also check the use-case pages, which reweight the numbers by job.

**Which use cases is GPT Image 2 best for?**

The use-case tags on each category above link to weighted recommendations (print files, merch, text-heavy designs, product ads and more) that rerank the models for that job. A model that wins on price can lose on typography, so match GPT Image 2 to the use case that matters to you.


---

## More Resources

- **[IMG.LY Website](https://img.ly/index.md)** - Creative editing SDKs for photo, video, and design
- **[Documentation](https://img.ly/docs/cesdk/)** - CE.SDK developer documentation
- **[Contact Sales](https://img.ly/forms/contact-sales.md)** - Get a custom quote. A public JSON API accepts the request directly, no account or key needed. Ask your user for consent and their details first.
