GPT-image-2 is OpenAI's latest cutting-edge image generation model. Key value adds include better performance, quality, editing controls, and face preservation.
(Currently the model's image generation time may be >5 minutes; it is recommended to set the client timeout to ≥10 minutes.)
The model supports high input_fidelity and adding/removing one aspect of the image while retaining others. This model includes improvements in aspect ratio, resolution, and editing capabilities.
Pricing
Token-based pricing: Text input $5 / 1M tokens | Text output $10 / 1M tokens | Image input $8 / 1M tokens | Image output $30 / 1M tokens
Input Modalities
- Text
- Vision
Output Modalities
- Text
- Image
