Multi-element creative briefs
Describe several subjects, their relationships, and the camera setup in one prompt — useful for storyboards, campaign key visuals, and scene-heavy concept work.
OpenAI text-to-image · Complex prompts · In-image typography
OpenAI's newest image model for complex prompts, in-image typography, and photoreal product work — text-to-image or edit with up to 16 references.
OpenAI’s newest image model for creators who write detailed prompts. GPT Image 2 follows long briefs reliably, renders legible text inside the frame, and produces convincing product and lifestyle photography — generate from scratch or edit with up to 16 reference images.
Create with GPT Image 2Describe several subjects, their relationships, and the camera setup in one prompt — useful for storyboards, campaign key visuals, and scene-heavy concept work.
Draft posters, packaging mocks, menu boards, and social ads where the words on the image need to be spelled correctly, not guessed.
Generate hero shots and brand scenes from text, or upload references and describe the retouch — background swaps, label updates, and multi-image composites across up to 16 inputs.

Same OpenAI family as ChatGPT image generation and the successor to DALL·E — with tighter prompt adherence, safer defaults, and a rendering pipeline tuned for production work.
“Create a polished campaign key visual from this complete art-direction brief while preserving the hierarchy and product silhouette.”

Long shot lists, crowded scenes, and layered instructions usually survive the full generation. Fewer missing props, fewer broken layouts when your brief runs past a paragraph.
“A busy design studio with exactly five people, two paper prototypes, one red task lamp, and a rainy city visible through the windows.”

Headlines, prices, labels, and body lines land with correct spelling and sensible line breaks — useful for posters, packaging, menus, covers, and slide-style layouts.
“A premium tea package mockup with the exact label ‘MIST GARDEN’ and the small line ‘SPRING HARVEST’. No other text.”
Skin, fabric, food, metal, and glass look believable without stacking modifier keywords — a strong default for e-commerce hero shots, brand campaigns, and editorial-style portraits.
Start with the subject, then add lighting, style, and anything that must stay the same. Upload a reference whenever you want a more controlled edit.
Lead with the subject, then add setting, lighting, lens, and a checklist of objects that must appear. GPT Image 2 is designed for prompts that would overwhelm lighter models.
Put headlines, prices, and labels in “quotes” and note placement — centered title, price tag bottom-right, small legal line along the footer.
For product and portrait work, specify material and lighting — brushed steel, matte glaze, soft window light from camera-left — instead of relying on generic “realistic” modifiers.
When uploading several images, say what each one contributes — “product from image 1, backdrop from image 2, lighting mood from image 3.” Supports up to 16 references.
Use medium quality to explore composition cheaply, then move to high quality and larger aspect ratios — including experimental 4K sizes — once the direction is locked.
Pick the version that fits your workflow — then open its page to create with that model locked in.
| Feature | GPT Image 2Current | ||
|---|---|---|---|
| Best for | Fast iteration through premium final creative | Complex prompts, in-image text, photoreal products | Everyday production work and cutouts |
| Speed | Fastest GPT Image generation with Flare | Balanced | Fast |
| On-image text | Best in the GPT Image family | Excellent | Very good |
| Transparent background | Supported | Not supported | Supported |
| Sizes & framing | Ratios and explicit sizes up to 4K | Wide range, up to 4K experimental | Standard ratios |
| Editing | Highest precision and reference fidelity | Up to 16 references | Strong prompt following + input fidelity |
Model basics, how it compares to ChatGPT and DALL·E, and practical tips for this page.
Explore popular image models outside this series, or browse the full catalog.




Bring a detailed brief, reference images, or an existing composition and create a refined result with stronger text, structure, and visual control.