Text-to-Image Generation
Generates images from text prompts at resolutions up to 1536×1024 and 1024×1536 pixels. Supports a 4000-token context window for detailed prompt input.
GPT Image 1.5 is an image generation model developed by OpenAI and released in December 2025. It serves as the engine behind the ChatGPT Images experience and is also available to developers via the API using the model ID gpt-image-1.5. The model supports both text-to-image generation and image editing, with output resolutions up to 1536×1024 and 1024×1536 pixels. It ranks first in the Image Edit category on the Chatbot Arena leaderboard. The model is designed to follow nuanced editing instructions, changing only the specified elements of an image while preserving lighting, composition, and overall context. It maintains consistent facial likeness across inputs, outputs, and sequential edits, making it well-suited for photo retouching, virtual try-ons, stylistic transformations, and multi-image compositing. Text rendering accuracy is notably improved compared to its predecessor, and generation speed is up to 4x faster. It is accessible to all ChatGPT users as well as developers building applications in e-commerce, marketing, and creative tooling.
High-signal model metadata in a structured two-column overview table.
The entity that provides this model.
The number of tokens supported by the input context window.
The number of tokens that can be generated by the model in a single request.
Whether the model's code is available for public use.
When the model was first released.
When the model's knowledge was last updated.
The providers that offer this model. This is not an exhaustive list.
Types of data this model can process.
A fuller summary of positioning, capabilities, and source-specific details for GPT Image 1.5.
GPT Image 1.5 is an image generation model developed by OpenAI and released in December 2025. It serves as the engine behind the ChatGPT Images experience and is also available to developers via the API using the model ID gpt-image-1.5. The model supports both text-to-image generation and image editing, with output resolutions up to 1536×1024 and 1024×1536 pixels. It ranks first in the Image Edit category on the Chatbot Arena leaderboard.
The model is designed to follow nuanced editing instructions, changing only the specified elements of an image while preserving lighting, composition, and overall context. It maintains consistent facial likeness across inputs, outputs, and sequential edits, making it well-suited for photo retouching, virtual try-ons, stylistic transformations, and multi-image compositing. Text rendering accuracy is notably improved compared to its predecessor, and generation speed is up to 4x faster. It is accessible to all ChatGPT users as well as developers building applications in e-commerce, marketing, and creative tooling.
Generates images from text prompts at resolutions up to 1536×1024 and 1024×1536 pixels. Supports a 4000-token context window for detailed prompt input.
Adds, removes, blends, or transposes elements in existing photos while preserving lighting and composition. Follows nuanced instructions to change only the specified parts of an image.
Maintains consistent facial likeness across inputs, outputs, and sequential edits. Useful for portrait retouching, virtual try-ons, and character consistency across a chain of edits.
Modifies specific regions of an image without affecting surrounding areas. Allows targeted edits such as replacing backgrounds or altering individual objects.
Generates images containing readable text with higher accuracy than the predecessor model. Suitable for creating mockups, signage, and branded visuals.
Blends subjects from multiple source images into a single coherent scene. Accepts image URL arrays as input to support multi-source compositing workflows.
Produces images up to 4x faster than its predecessor model. Speed improvements apply to both generation and editing tasks.
Available via the OpenAI API using the model ID gpt-image-1.5. Suitable for developers integrating image generation into e-commerce, marketing, and creative applications.
Primary API pricing shown in the same “quick compare” spirit as the reference page.
Additional usage-cost dimensions synced into the project for this model.
Places where this model is available, based on the synced detail-page metadata.
The configurable options currently documented for this model.
If you want to edit an existing image, provide the URL(s) or variables
Parameters currently listed by OpenRouter or the local catalog for this model.
Official model cards, release notes, docs, and other references synced from the source page.
Recent daily stories tied to GPT Image 1.5 through direct model mentions or provider-level coverage.
Anthropic and OpenAI move deeper into real workflows.
Anthropic and Hugging Face move deeper into real workflows.
NVIDIA and Hugging Face move deeper into real workflows.
OpenAI and Hugging Face move deeper into real workflows.
GPT Image 1.5 discussions are most active in r/ChatGPT, r/singularity, r/OpenAI. Top Reddit threads cluster around benchmark and model-comparison threads, coding workflow discussions.
The strongest match in this snapshot has 1157 upvotes and 240 comments.
The image generation war just heated up again. OpenAI has officially dropped **GPT-Image-1.5** and it has already dethroned Google on the leaderboards.
**The Benchmarks (LMArena):**
**Rank:** #1 Overall in Text-to-Image With **Score** 1277 (Beating Gemini 3 Pro Image / Nano Banana Pro at 1235).
**Key Upgrades:**
**Speed:** 4x Faster than the previous model (DALL-E 3 / GPT-Image-1).
**Editing:** It supports precise "add, subtract, combine" editing instructions.
**Consistency:** Keeps character appearance and lighting consistent across edits (a major pain point in DALL-E 3).
**Availability:** ChatGPT: Rolling out today to all users via a new "Images" tab in the sidebar.
**API:** Available immediately as gpt-image-1.5.
**Google held the crown with "Nano Banana Pro" for about a month. With OpenAI claiming "4x speed" and better instruction following, is this the DALL-E 3 successor we were waiting for?**
**Source: OpenAI Blog**
🔗: https://openai.com/index/new-chatgpt-images-is-here/
**Video :** https://youtu.be/DPBtd57p5Mg?si=iBlvJ0Km6uUoltYn
Many people did not like my "realistic" results, so i tried again. Still not perfect, but better than before.
The first 3 images of each set are GPT image 1.5, the rest is Nano Banana Pro.
I think Nano Banana Pro won this round.
Introducing ChatGPT Images, powered by our flagship new image generation model.
* Stronger instruction following
* Precise editing
* Detail preservation
* 4x faster than before
Rolling out today in ChatGPT for all users, and in the API as GPT-Image-1.5.
[https://openai.com/index/new-chatgpt-images-is-here/](https://openai.com/index/new-chatgpt-images-is-here/)
Have seen a lot of examples from both models and I can say pretty surely that nana banana pro is much better than gpt-image-1.5.
What do you guys think?
GPT Image 1.5 supports a context window of 4000 tokens, which is used for processing text prompt input.
The model is accessible through the OpenAI API using the model ID gpt-image-1.5. It accepts text prompts and image URL arrays as inputs, making it suitable for both generation and editing workflows.
The model supports output resolutions up to 1024×1536 pixels (portrait) and 1536×1024 pixels (landscape).
Based on the available metadata, GPT Image 1.5 has a training date of December 2025.
The model supports adding, removing, combining, blending, and transposing elements in existing images, as well as inpainting specific regions, face preservation across edits, and multi-image compositing from multiple source photos.
Yes, GPT Image 1.5 is available directly in ChatGPT for all users and also accessible via the OpenAI API for developers.
Continue browsing adjacent models from the same provider.