01
**Precise in-image text rendering**. blends readable symbols, typography, and labels directly into generated imagery, making it practical for logos, posters, and infographics
02
**Strong instruction following**. handles up to 10–20 distinct objects in a single scene with accurate trait and spatial binding, well beyond what most image models manage
03
**Image API: Generations endpoint**. creates images from a text prompt with no reference required
04
**Image API: Edits endpoint**. modifies existing images partially or entirely using a new prompt, replacing or extending specific regions
05
**Responses API: multi-turn editing**. supports iterative high-fidelity edits via conversation context. Accepts image File IDs as input. Action mode can be set to auto, generate, or edit
06
**In-context learning from reference images**. analyzes uploaded images and integrates their details into generation, enabling style transfer and scene continuation
07
**Output controls**. adjust quality, size, format, compression, and transparent background support. Generate multiple images in one request via the `n` parameter
08
**Wide style range**. photorealism, vintage film photography, Polaroid, 3D rendering, concept art, product mockups, infographics, and diagrams
09
**C2PA provenance metadata**. every output is tagged as AI-generated by GPT-4o, giving buyers and publishers a verifiable content credential