Generate high-quality images from text prompts using GPT Image 2, supporting up to 20,000 character prompts for detailed and precise image creation.
About this model
GPT Image 2 Text to Image generates high-quality images from detailed text prompts, supporting up to 20,000 characters for precise creative control. It delivers photorealistic results, accurate text rendering, and strong instruction-following for complex scenes and compositions, with selectable quality (low / medium / high) and 1K / 2K / 4K resolutions.
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.025 – $0.150 per image | Pay only for what you use. Low: $0.025 / $0.040 / $0.075 (1K / 2K / 4K). Medium: $0.030 / $0.045 / $0.090. High: $0.060 / $0.090 / $0.150. |
| Fal.ai | $0.040 – $0.110 per image | Low and medium quality only. Low/Medium: $0.040 / $0.060 / $0.110 (1K / 2K / 4K). |
| Replicate | Not available | GPT Image 2 is not available on Replicate. |
Pay only for what you use. Low: $0.025 / $0.040 / $0.075 (1K / 2K / 4K). Medium: $0.030 / $0.045 / $0.090. High: $0.060 / $0.090 / $0.150.
Low and medium quality only. Low/Medium: $0.040 / $0.060 / $0.110 (1K / 2K / 4K).
GPT Image 2 is not available on Replicate.
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| Prompt | string | Text prompt describing the image to generate. Maximum 20,000 characters. | A girl doing livestream makeup tutorials while tiny workers inside her mirror physically repaint the reflection in real time, miniature ladders, paint buckets, glowing vanity lights, cluttered modern bedroom aesthetic, insanely detailed reflections, playful surreal realism, highly creative composition, luxury commercial quality. |
| Aspect Ratio | Enum (10 options) | Output image aspect ratio. Note: images with aspect ratio 'auto' (or unspecified) will only be converted to 1K; 1:1 cannot be converted to 4K — otherwise the task will fail to create. | auto |
| Resolution | Enum (3 options) | Image resolution. Note: images with a 1:1 aspect ratio cannot be converted to 4K. Images with aspect ratio 'auto' (or unspecified) will only be converted to 1K; otherwise the task will fail to create. | 2K |
| Quality | Enum (3 options) | Generation quality. 'low' and 'medium' use a faster, cheaper backend; 'high' uses the higher-fidelity backend. | high |
Text prompt describing the image to generate. Maximum 20,000 characters.
A girl doing livestream makeup tutorials while tiny workers inside her mirror physically repaint the reflection in real time, miniature ladders, paint buckets, glowing vanity lights, cluttered modern bedroom aesthetic, insanely detailed reflections, playful surreal realism, highly creative composition, luxury commercial quality.Output image aspect ratio. Note: images with aspect ratio 'auto' (or unspecified) will only be converted to 1K; 1:1 cannot be converted to 4K — otherwise the task will fail to create.
autoImage resolution. Note: images with a 1:1 aspect ratio cannot be converted to 4K. Images with aspect ratio 'auto' (or unspecified) will only be converted to 1K; otherwise the task will fail to create.
2KGeneration quality. 'low' and 'medium' use a faster, cheaper backend; 'high' uses the higher-fidelity backend.
highDeveloper documentation
Write your prompt: Describe the image you want in detail. You can include style, lighting, subject, background, and mood. Up to 20,000 characters are supported.
Pick a quality: Choose low or medium for fast, lower-cost drafts, or high for the highest fidelity output. high is the default.
Pick aspect ratio and resolution: Select an aspect ratio (e.g. 1:1, 16:9, 9:16) and a resolution (1K, 2K, or 4K). Note that 1:1 cannot be rendered at 4K, and auto aspect ratio is locked to 1K.
Submit the request: Click Generate. The model processes your prompt and returns the generated image.
Review and iterate: If the result needs adjustment, refine your prompt with more specific instructions and regenerate.
Frequently asked
Be specific about the subject, style, lighting, background, and mood. More detail generally produces better results. Prompts up to 20,000 characters are supported.
`low` and `medium` are faster and cheaper and are well suited to drafts and bulk generation. `high` runs the full-fidelity model and is best for finals. Pricing scales with both quality and resolution.
Each request generates one image. Submit multiple requests to generate variations.
The model supports a wide range of styles including photorealistic, illustration, concept art, product photography, and more.