GPT Image (OpenAI)
OpenAI's image model — DALL-E's successor in ChatGPT and the API
Overview
GPT Image is OpenAI's current image generation model, accessible through ChatGPT and the API. OpenAI retired DALL-E 2 and DALL-E 3 on May 12, 2026; the current flagship, GPT Image 2 (released April 21, 2026, also branded "ChatGPT Images 2.0"), is architecturally different from its DALL-E predecessors — it's autoregressive (generating images token-by-token like text) rather than diffusion-based, with roughly 99% text-rendering accuracy (up from 90-95% on DALL-E 3), native 4K output, built-in reasoning/planning, and the ability to pull web information before generating.
GPT Image is integrated directly into ChatGPT Plus and ChatGPT Enterprise — users generate images conversationally, asking ChatGPT to refine results through natural language — and is also available via the OpenAI image API for developers (the older DALL-E API endpoints are deprecated). The model refuses to generate content that violates OpenAI's policies, including realistic depictions of public figures and certain sensitive content, and declines to reproduce artists' styles if they've opted out of the training dataset.
Pricing
Pricing shown for reference only. These figures reflect RECATOOLS research as of 11 Jul 2026 and may be out of date or incomplete. This is not financial or purchasing advice — always confirm the current price on the provider’s official website before making any decision.
Use cases
ASEAN Perspective
GPT Image (OpenAI) in Southeast Asia
ASEAN-region availability and pricing notes coming soon. Drop the editorial team a note via /contact/ if you can supply local context (Singapore/Malaysia/Indonesia/Thailand/Vietnam).
GPT Image 2 is OpenAI's current image model — DALL-E 2 and 3 were retired May 12, 2026 — and its headline win is text: multi-line copy and even CJK scripts render reliably, which no mainstream rival matches. It reasons before generating, can pull web references mid-task, and outputs up to 4K, though OpenAI itself calls anything above 2K experimental. ChatGPT integration makes iteration conversational, and API pricing is usage-based (roughly $0.21 for a high-quality 1024px image, $0.41 at 4K), so production volume adds up fast. Fine artistic control still trails Midjourney, designers complain natural language is a blunt instrument for precise layout work, and content filters are strict. For on-brief images with legible text it's the default choice; ASEAN users get broad prompt-language tolerance and solid API documentation.
What people say
May 12, 2026 was the day DALL-E actually died: OpenAI cut off API access to DALL-E 2 and 3, finishing a migration that began when GPT Image 1.5 quietly replaced DALL-E 3 as ChatGPT's default in December 2025. The current model, GPT Image 2 (launched April 21, 2026, branded ChatGPT Images 2.0 in the app), is a different animal — autoregressive rather than diffusion-based, reasoning before it generates, and able to search the web mid-task for reference material.
User reaction centres on two things. Text rendering is the breakout: reviewers report multi-line copy, mixed font weights, and even CJK scripts holding up, and an r/ChatGPTPro thread called the character-consistency and text upgrades "wild." The second is the look — one r/LocalLLM user credited it with finally killing the yellow-tinted AI-art filter. Complaints cluster around control and cost. Graphic designers call natural-language steering "terrible for graphic design specifically" since there's no way to pin exact positions, and API pricing runs about $0.006, $0.053, and $0.211 per 1024px image at low, medium, and high quality — $0.41 for a high-quality 4K frame, which stings at production volume. OpenAI itself concedes resolutions above 2K are experimental and inconsistent; treat 2K as the reliable ceiling.
The forced migration annoyed developers — an OpenAI community thread titled "OpenAI is making a huge mistake by deprecating DALL-E 3" gathered teams whose pipelines depended on the older model's cheaper, faster responses. For most users, though, the trade is straightforwardly favourable: this is the first mainstream model where in-image text just works.
Summary of public user & expert reviews, compiled by RECATOOLS.
Notable facts
- DALL-E 3 was trained with a technique where human annotators provided extremely detailed captions for every training image, which is why it follows complex prompts more accurately than any predecessor.
- The model can accurately count objects in a scene — generating exactly 7 red apples on a table, for example — something that defeated every previous AI image generator.
- DALL-E 3's content policy rejects requests to generate images of real living people by name, a deliberate choice to prevent misuse for misinformation.
Frequently asked questions
About this listing
This entry was compiled from publicly available data including GPT Image (OpenAI)'s official website, press releases, documentation, and reputable third-party publications. RECATOOLS is not affiliated with GPT Image (OpenAI) unless explicitly stated.
Third-party AI tools update their pricing, features, availability, and policies frequently. Information here may be outdated by the time you read this — we make reasonable efforts to keep listings current, but cannot guarantee absolute accuracy.
For the latest details, please refer to GPT Image (OpenAI) directly →
Spotted something out of date? Suggest an update →
GPT Image (OpenAI) in the news
AI & ML
AWS Commits $1 Billion to Embedded AI Engineers as the Enterprise Fight Shifts to Deployme...
AI & ML
OpenAI Ships GPT-5.6 to Everyone — Three Tiers, One Generation, and a Government Preview I...
AI & ML
OpenAI Reportedly Floats a 5% US Government Stake — and Reopens the Question of Public Own...
More in Image Generation