OpenAI released ChatGPT Images 2.0 on April 21, and the feature most reviewers are pointing to is not the resolution increase or the broader style range. It is the reasoning pass that now runs before the model generates anything.

Earlier image generators, including OpenAI's own DALL-E line, translated a prompt directly into pixels. The new model, gpt-image-2, runs a reasoning step first: it searches the web for layout references where helpful, plans a composition, and checks the output against the original instructions before returning a result. The practical difference is that detailed instructions now land more reliably. Object placement, spatial relationships, and coherence across a set of related images all improve in proportion to how specific the prompt is.

Two tiers

The model ships in two modes. Instant mode is available to all users, including the free tier, and delivers the core quality improvements over DALL-E 3 without the reasoning overhead. Thinking mode, restricted to Plus, Pro, Business, and Enterprise subscribers, enables the web search, layout reasoning, multi-image batching of up to eight images per prompt, and an output verification step.

The gap between tiers is real. Generating a matched set of product images that share backgrounds, angles, and lighting requires the thinking pass. For a single illustration, the free tier will generally do. TechCrunch and MacRumors both flagged the text rendering improvements as the most immediately visible change, particularly for non-Latin scripts including Japanese, Korean, Hindi, and Bengali, where previous models produced convincing-looking but incorrect characters.

What it replaces

DALL-E 2 and DALL-E 3 are being retired on May 12, 2026. Images 2.0 is not a supplement to the existing line; it is the replacement, available to all ChatGPT, Codex, and API users. OpenAI also released Codex Labs the same day, a developer training service tied to its Codex platform, continuing its push to move AI tools deeper into production workflows rather than keeping them as standalone assistants.

The model outputs at up to 2K resolution and generates up to eight coherent visuals from a single prompt. OpenAI is clearly trying to close the gap between AI-generated assets and the kind of polished deliverables that currently require a designer and dedicated software. How far that gap actually closes in practice is something studios and agencies will spend the next few months finding out.

Sources

  1. i. openai.com
  2. ii. techcrunch.com
  3. iii. www.macrumors.com
  4. iv. www.axios.com

Commentarii · 0

Add · a · Comment