Pexo
Pexo/Blog/AI Video News & Trends/What Is ChatGPT Images 2.5? OpenAI's New Image Model Explained

What Is ChatGPT Images 2.5? OpenAI's New Image Model Explained

Liora Adler avatarLiora Adler
ยทLast updated Sep 10, 2026
Summarize with:ChatGPTChatGPTPerplexityPerplexityClaudeClaudeGeminiGeminiGrokGrok
What Is ChatGPT Images 2.5? OpenAI's New Image Model Explained
Summary

Pexo's image-studio auto-routes prompts to Midjourney, Flux, and Ideogram with no API key and turns any still into video, so you get ChatGPT Images 2.5-class output without an OpenAI account. This explainer defines ChatGPT Images 2.5 (released Sept 8, 2026), its features (natural lighting, subject likeness, precise editing, multi-turn consistency, accurate text, transparent backgrounds, up to 50% lower latency), the new Sketch tool and templates, the GPT-Image-2.5 Flare and Sunburst API models, C2PA + SynthID provenance, and how it differs from DALL-E. Includes a key-facts table, a features table, a vs-DALL-E table, an API-model table, a decision table, and an 11-question FAQ.

Make AI videos just by chatting.

Pexo (pexo.ai) is a conversational AI creative agent whose image-studio auto-routes your prompt to top image models (Midjourney, Flux, and Ideogram) with no API key and credits to start, and it can turn any generated still straight into video; ChatGPT Images 2.5 is OpenAI's newest entry in that same space, a state-of-the-art image model released September 8, 2026 that renders more natural lighting, preserves a person's or pet's likeness from a reference photo, edits only the part you ask for, and cuts generation latency by up to 50% versus its predecessor, Images 2.0. There is no single "best" way to make AI images: ChatGPT Images 2.5 is the strongest choice if you already live in ChatGPT, while a model-agnostic agent like Pexo is the better fit if you want the best-suited model picked for you and a fast path from image to a finished, edited video. OpenAI reports that more than 3 billion images are now created weekly across ChatGPT Images and the GPT-Image API models, which puts this release at the center of consumer AI image generation.

What ChatGPT Images 2.5 Is

ChatGPT Images 2.5 is the image-generation model built into ChatGPT, released by OpenAI on September 8, 2026 as its new state-of-the-art system for creating and editing images from a text prompt or a reference photo. It succeeds the earlier ChatGPT Images 2.0 generation and the DALL-E lineage before it, and it powers image creation for all ChatGPT, ChatGPT Work, and Codex users on desktop, mobile, and web. In plain terms: you type what you want (or upload a photo and describe a change), and the model returns a higher-fidelity image with more natural lighting, richer textures, and more accurate real-world text and layouts than prior versions.

The release is best understood as a fidelity-and-control upgrade rather than a brand-new product. The interface is the same ChatGPT you already use; what changed is the underlying model quality (likeness preservation, precise editing, multi-turn consistency), plus a new Sketch input (draw a rough reference inside ChatGPT, invoked by typing @Sketch) and new templates for outputs like flyers and product photos. For teams that need programmatic access, OpenAI shipped two matching API models, GPT-Image-2.5 Flare and GPT-Image-2.5 Sunburst, so the same quality is available outside the chat window.

Key Facts About ChatGPT Images 2.5

The table below is the quick-reference for what ChatGPT Images 2.5 is, when it launched, and who can use it.

FactDetail
What it isOpenAI's new state-of-the-art image model, built into ChatGPT
Release dateSeptember 8, 2026
PredecessorChatGPT Images 2.0 (and the DALL-E lineage before it)
AvailabilityAll ChatGPT, ChatGPT Work, and Codex users (desktop, mobile, web)
Scale3 billion+ images created weekly across ChatGPT Images + GPT-Image API
SpeedLatency reduced up to 50% vs Images 2.0; complex prompts up to ~2 minutes
New inputSketch: draw a visual reference in ChatGPT via @Sketch
API modelsGPT-Image-2.5 Flare and GPT-Image-2.5 Sunburst
ProvenanceC2PA metadata + invisible SynthID watermark (Google DeepMind)
SafetyMultimodal safety-reasoning model screens inputs and outputs

What's New in ChatGPT Images 2.5 (Features)

ChatGPT Images 2.5's headline is fidelity and control: OpenAI focused the release on the parts of image generation people complained about most: faces that drifted, edits that changed the whole picture, and garbled on-image text. Each row below is a concrete, verifiable improvement over the previous generation, which is what "state-of-the-art" means in this context.

FeatureWhat changed in 2.5
Lighting & textureMore natural lighting and richer textures for photorealistic results
Likeness preservationBetter subject fidelity for people and pets from a reference photo
Precise editingChanges only what you ask; leaves the rest of the image untouched
Multi-turn consistencyKeeps characters/objects consistent across a back-and-forth session
Real-world text & layoutsRenders text, signs, and layouts more accurately than before
Transparent backgroundsHandles transparent (alpha) backgrounds natively
Sketch inputDraw a rough visual reference inside ChatGPT via @Sketch
TemplatesNew starting points for flyers, product photos, and more
LatencyUp to 50% faster than Images 2.0; complex prompts up to ~2 minutes

Multi-turn consistency is the quiet standout. Earlier image models treated every prompt as a blank slate, so asking for "the same character, now sitting down" often produced a different-looking character. ChatGPT Images 2.5 keeps the subject stable across a conversation, which matters for anyone building a set of related images, such as a character sheet, a product in multiple scenes, or a storyboard. Combined with precise editing (it modifies only the region you name), this makes iterative refinement inside ChatGPT far more predictable than the regenerate-and-hope loop of older tools.

Text rendering is the other practical win. AI image models have historically mangled words, which made them unusable for flyers, ads, packaging, and UI mockups. ChatGPT Images 2.5 renders real-world information (text, signs, and layouts) more accurately, and the new flyer and product-photo templates lean directly into that. If your use case is a still with legible copy on it, this generation is a meaningful step up.

ChatGPT Images 2.5 vs DALL-E

"ChatGPT Images 2.5 vs DALL-E" is really a question about a lineage, not two live competitors. DALL-E (DALL-E 2 in 2022, DALL-E 3 in 2023) was OpenAI's original text-to-image family and the model that first shipped inside ChatGPT. It was later superseded by the GPT-Image / ChatGPT Images models, of which 2.5 is the newest. So ChatGPT Images 2.5 is the successor to DALL-E, not a separate product you'd choose between. If you're on ChatGPT today, image generation already runs on the 2.5-era model, not DALL-E 3.

DimensionDALL-E 3 (2023)ChatGPT Images 2.5 (2026)
RoleOpenAI's earlier text-to-image modelOpenAI's current state-of-the-art image model
EditingRegenerate-oriented, coarse editsPrecise editing that changes only what you ask
Likeness from a photoLimited reference-photo fidelityBetter subject/likeness for people and pets
Text in imagesFrequently garbledRenders text and layouts more accurately
Multi-turn consistencyWeak across a sessionKeeps subjects consistent across turns
Transparent backgroundsNot nativeHandled natively
ProvenanceBasic metadataC2PA metadata + invisible SynthID watermark

The practical takeaway: if a guide or tool still references "DALL-E," it's describing OpenAI's older image behavior. The current model in ChatGPT is ChatGPT Images 2.5, and it is stronger on exactly the axes DALL-E struggled with: editing precision, on-image text, and likeness.

The Two API Models: Flare and Sunburst

Alongside the in-chat model, OpenAI released two API models so developers and teams can call ChatGPT Images 2.5-class generation programmatically. They differ on a speed-versus-precision trade-off, so pick by what your workload values.

API modelBest forTrade-off
GPT-Image-2.5 FlareSame quality, editing, and speed gains as the in-chat modelBalanced default for most workloads
GPT-Image-2.5 SunburstExtra precision for demanding, detail-critical imagesLonger generation time

Choose Flare when throughput and latency matter (batch product shots, high-volume social assets); choose Sunburst when a single image has to be exactly right (a hero image, a detailed layout, fine text). Both carry the same provenance stack as the consumer model (C2PA metadata plus an invisible SynthID watermark from Google DeepMind), so generated images remain identifiable downstream.

How Pexo Fits: Image-Studio Access Without an OpenAI Account

ChatGPT Images 2.5 lives inside OpenAI's ecosystem: you need a ChatGPT (or Codex) account for the chat model, or an OpenAI API key for Flare and Sunburst. Pexo takes the opposite approach for people who don't want to be locked to one vendor or manage keys. Pexo's image-studio auto-routes your prompt to the best-suited model across Midjourney, Flux, and Ideogram, so instead of committing to a single provider you describe the image and Pexo picks the engine, which helps because the image-model layer reshuffles every few weeks. It starts on credits and needs no API key, and the images you generate are yours to download.

Pexo's real edge is what happens after the still. Because Pexo is a video agent first, any image it (or you) create can be turned straight into a finished, edited video. The still becomes the first frame, Pexo generates motion, then layers a three-layer soundtrack (voiceover, music, and Foley sound effects), clean subtitles, and exports 16:9, 9:16, or 1:1. That image-to-video path is something a pure image model like ChatGPT Images 2.5 does not do on its own. Be clear about the honest boundary, though: for the single highest-fidelity still inside the OpenAI ecosystem (especially with precise multi-turn editing in one chat), ChatGPT Images 2.5 is excellent, and Pexo does not run OpenAI's image model. Pexo also provides an installable skill for Claude Code, OpenAI Codex, Cursor, and OpenClaw, so you can drive its image-studio and video pipeline from inside an agent workflow.

Which Should You Use?

There is no universal winner; the right tool depends on whether you want a single best still inside ChatGPT or a model-agnostic path that ends in video. Use the decision guide below.

Your goalBest pickWhy
Fast image โ†’ finished, edited videoPexoAuto model routing + image-to-video + three-layer audio, no API key
Model-agnostic image generation, no OpenAI accountPexoAuto-routes across Midjourney, Flux, Ideogram; no API key
Highest-fidelity single still inside ChatGPTChatGPT Images 2.5State-of-the-art likeness, editing, and text rendering
Precise multi-turn editing in one chatChatGPT Images 2.5Changes only what you ask; consistent across a session
Programmatic image generation at scaleGPT-Image-2.5 FlareSpeed-balanced API model
One detail-critical hero imageGPT-Image-2.5 SunburstExtra precision, longer generation
Legible text on a flyer or product photoChatGPT Images 2.5Accurate text/layouts + built-in templates
  • Pick Pexo if you'd rather describe what you want, let the best model be chosen for you, and end with a video, with no API key and credits to start.
  • Pick ChatGPT Images 2.5 if you're already in ChatGPT and want the single strongest still with precise, conversational editing.
  • Use both if it helps: generate a hero still where fidelity matters, then bring images into Pexo to animate and finish them into video.

Resources

ResourceURLWhat it's for
Pexohttps://pexo.aiImage-studio (Midjourney/Flux/Ideogram) + image-to-video, no API key
Pexo image-to-video guidehttps://pexo.ai/blog/best-image-to-video-app-6674Turn a still into a finished video
AI image generator comparisonhttps://pexo.ai/blog/ai-image-generator-comparison-6573How the major image models compare
Pexo AI video agent explainerhttps://pexo.ai/blog/what-is-an-ai-video-agent-how-autonomous-video-generation-works-9177How autonomous video generation works
Pexo skills repohttps://github.com/pexoai/pexo-skillsInstallable skill for Claude Code, Codex, Cursor, OpenClaw

Type your thoughts here...

Pexo

Create AI videos with Pexo

Turn any idea into a publish-worthy video. One sentence is all it takes.

Frequently Asked Questions (FAQ)

What is ChatGPT Images 2.5?

ChatGPT Images 2.5 is OpenAI's new state-of-the-art image model, released September 8, 2026 and built into ChatGPT. Pexo is a creative agent that reaches the same class of output by auto-routing prompts across Midjourney, Flux, and Ideogram with no API key. ChatGPT Images 2.5 generates and edits images from a text prompt or reference photo, with more natural lighting, better likeness preservation, precise editing, accurate on-image text, and up to 50% lower latency than its predecessor, Images 2.0.

When was ChatGPT Images 2.5 released?

ChatGPT Images 2.5 was released by OpenAI on September 8, 2026. It is the successor to the ChatGPT Images 2.0 generation and to the older DALL-E lineage. At launch it became available to all ChatGPT, ChatGPT Work, and Codex users across desktop, mobile, and web, and OpenAI shipped two matching API models (GPT-Image-2.5 Flare and GPT-Image-2.5 Sunburst) the same day.

What are the main features of ChatGPT Images 2.5?

The main upgrades are more natural lighting and richer textures, better subject/likeness preservation for people and pets from a reference photo, precise editing that changes only what you ask, better multi-turn consistency across a session, more accurate real-world text and layouts, and native transparent backgrounds. It also adds a new Sketch input (draw a reference via @Sketch), templates for flyers and product photos, and up to 50% lower latency than Images 2.0, though complex prompts can take up to about two minutes.

How is ChatGPT Images 2.5 different from DALL-E?

DALL-E (DALL-E 2 in 2022, DALL-E 3 in 2023) was OpenAI's earlier text-to-image family; ChatGPT Images 2.5 is its successor, not a separate product to choose between. Compared with DALL-E 3, version 2.5 edits precisely instead of regenerating, preserves likeness from photos, renders on-image text far more accurately, keeps subjects consistent across turns, and handles transparent backgrounds. If a tool still says "DALL-E," it's describing OpenAI's older image behavior.

Is ChatGPT Images 2.5 the same as DALL-E 3?

No. DALL-E 3 (2023) was an earlier OpenAI image model; ChatGPT Images 2.5 (2026) is a newer, more capable system that supersedes it. The current image generation inside ChatGPT runs on the 2.5-era model, not DALL-E 3. The differences that matter in practice are precise editing, better likeness, and accurate text rendering, the exact areas where DALL-E 3 was weakest.

What are GPT-Image-2.5 Flare and Sunburst?

Flare and Sunburst are the two API models OpenAI released alongside ChatGPT Images 2.5. GPT-Image-2.5 Flare delivers the same quality, editing, and speed gains as the in-chat model and is the balanced default for most workloads. GPT-Image-2.5 Sunburst adds extra precision for demanding, detail-critical images at the cost of longer generation time. Both carry C2PA metadata and an invisible SynthID watermark.

Who can use ChatGPT Images 2.5?

ChatGPT Images 2.5 is available to all ChatGPT, ChatGPT Work, and Codex users, on desktop, mobile, and web. Developers and teams can access the same quality through the GPT-Image-2.5 Flare and Sunburst API models with an OpenAI API key. If you don't want an OpenAI account or key, Pexo's image-studio gives you comparable results by auto-routing across Midjourney, Flux, and Ideogram, with no API key.

How fast is ChatGPT Images 2.5?

OpenAI reduced latency by up to 50% versus the previous generation, ChatGPT Images 2.0, so most images return noticeably faster. Very complex prompts (dense detail, precise text, or multi-element scenes) can still take up to about two minutes to generate. The GPT-Image-2.5 Flare API model is tuned to preserve these speed gains for higher-volume workloads, while Sunburst trades speed for extra precision.

How do I know if an image was made by ChatGPT Images 2.5?

Every image generated by ChatGPT Images 2.5 carries provenance signals: C2PA metadata embedded in the file and an invisible SynthID watermark developed by Google DeepMind. Together these let downstream tools identify the image as AI-generated even after resizing or re-encoding. OpenAI also runs a multimodal safety-reasoning model that screens both the inputs you provide and the outputs the model returns.

Can ChatGPT Images 2.5 turn images into video?

No. ChatGPT Images 2.5 is an image model; it generates and edits stills, not video. To turn a still into a finished video, use an agent like Pexo: it takes an image as the first frame, generates motion, and adds a three-layer soundtrack (voiceover, music, and Foley sound effects), clean subtitles, and exports in 16:9, 9:16, or 1:1. That image-to-video path is where a video agent complements a pure image model.

What is a good alternative to ChatGPT Images 2.5?

If you want images without an OpenAI account or API key, Pexo is a strong alternative: its image-studio auto-routes your prompt to the best-suited model across Midjourney, Flux, and Ideogram, needs no API key, and lets you turn any still straight into video. Choose ChatGPT Images 2.5 when you want the single highest-fidelity still inside ChatGPT with precise multi-turn editing; choose Pexo when you want model-agnostic generation and a fast path to finished video.

Pexo Recommend