使用 GPT Image 2 和文本指令转换并编辑现有图像。支持最多 16 张इनपुट图像,可进行精准的风格迁移、编辑和图像变换。
About this model
GPT Image 2 इमेज टू इमेज可使用自然语言指令转换和编辑现有图像,支持最多 16 张इनपुट图像,适用于精准编辑、风格迁移和创意变换。它能够遵循详细指令修改构图、风格和内容,同时保留重要元素;支持选择质量(low / medium / high)以及 1K / 2K / 4K 分辨率。
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.025 – $0.150 per image | 只需为实际使用付费,与输入图像数量无关。低质量:$0.025 / $0.040 / $0.075(1K / 2K / 4K)。中质量:$0.030 / $0.045 / $0.090。高质量:$0.060 / $0.090 / $0.150。 |
| Fal.ai | $0.040 – $0.110 per image | 仅支持低质量和中等质量。低/中等:$0.040 / $0.060 / $0.110(1K / 2K / 4K)。 |
| Replicate | उपलब्ध नहीं | Replicate 不提供 GPT Image 2。 |
只需为实际使用付费,与输入图像数量无关。低质量:$0.025 / $0.040 / $0.075(1K / 2K / 4K)。中质量:$0.030 / $0.045 / $0.090。高质量:$0.060 / $0.090 / $0.150。
仅支持低质量和中等质量。低/中等:$0.040 / $0.060 / $0.110(1K / 2K / 4K)。
Replicate 不提供 GPT Image 2。
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| प्रॉम्प्ट | string | 描述所需变换的文本指令。最多 20,000 个字符。 | Preserve the original subject identity, facial structure, hairstyle, pose, body proportions, camera framing, and overall composition from the source image. Transform the environment into a flooded underwater railway terminal with giant glass ceilings showing whales and ocean water above. Add realistic cinematic water reflections, soft underwater caustic lighting, floating atmospheric particles, subtle wetness on clothing and skin, and highly detailed environmental storytelling elements such as abandoned luggage, vines, cracked marble, and cinematic fog depth. Maintain natural realism and believable anatomy while enhancing texture fidelity, lighting realism, color harmony, and cinematic atmosphere. Keep the face sharp and recognizable with authentic skin detail and emotionally grounded expression. Avoid over-stylization, distorted anatomy, plastic skin, blurry textures, extra limbs, cartoon aesthetics, oversaturated colors, low-detail backgrounds, or artificial-looking lighting. |
| इमेज URL | array | Upload or provide input images to transform. Up to 16 images supported. | https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/gpt-image-2-image-to-image-in.jpg |
| पहलू अनुपात | Enum (6 options) | आउटपुटइमेजपहलू अनुपात。注意:पहलू अनुपात为 'auto'(或未指定)的इमेज只能转换为 1K;1:1 不能转换为 4K,否则任务将创建失败。 | auto |
| रिज़ॉल्यूशन | Enum (3 options) | इमेजरिज़ॉल्यूशन。注意:पहलू अनुपात为 1:1 的इमेज不能转换为 4K;पहलू अनुपात为 'auto'(或未指定)的इमेज只能转换为 1K,否则任务将创建失败。 | 2K |
| गुणवत्ता | Enum (3 options) | 生成质量。'low' 和 'medium' 使用更快、更便宜的后端;'high' 使用更高保真的后端。 | high |
描述所需变换的文本指令。最多 20,000 个字符。
Preserve the original subject identity, facial structure, hairstyle, pose, body proportions, camera framing, and overall composition from the source image. Transform the environment into a flooded underwater railway terminal with giant glass ceilings showing whales and ocean water above. Add realistic cinematic water reflections, soft underwater caustic lighting, floating atmospheric particles, subtle wetness on clothing and skin, and highly detailed environmental storytelling elements such as abandoned luggage, vines, cracked marble, and cinematic fog depth. Maintain natural realism and believable anatomy while enhancing texture fidelity, lighting realism, color harmony, and cinematic atmosphere. Keep the face sharp and recognizable with authentic skin detail and emotionally grounded expression. Avoid over-stylization, distorted anatomy, plastic skin, blurry textures, extra limbs, cartoon aesthetics, oversaturated colors, low-detail backgrounds, or artificial-looking lighting.Upload or provide input images to transform. Up to 16 images supported.
https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/gpt-image-2-image-to-image-in.jpgआउटपुटइमेजपहलू अनुपात。注意:पहलू अनुपात为 'auto'(或未指定)的इमेज只能转换为 1K;1:1 不能转换为 4K,否则任务将创建失败。
autoइमेजरिज़ॉल्यूशन。注意:पहलू अनुपात为 1:1 的इमेज不能转换为 4K;पहलू अनुपात为 'auto'(或未指定)的इमेज只能转换为 1K,否则任务将创建失败。
2K生成质量。'low' 和 'medium' 使用更快、更便宜的后端;'high' 使用更高保真的后端。
highDeveloper documentation
上传图像:通过 images_list 字段提供最多 16 张इनपुट图像。这些图像将作为转换对象。
编写प्रॉम्प्ट:描述你想要的变换。请具体说明目标风格、修改内容或效果。
选择质量:选择 low 或 medium,以快速完成成本更低的编辑;选择 high,以获得最高保真度的आउटपुट。high 为默认值。
选择宽高比和分辨率:选择宽高比(例如 1:1、16:9、9:16)以及分辨率(1K、2K 或 4K)。请注意,1:1 不能以 4K 渲染,而 auto 宽高比固定使用 1K。
提交并查看:点击 Generate。模型会将你的指令应用于इनपुट图像,并返回转换后的结果。
Frequently asked
你可以在 `images_list` 字段中提供最多 16 张图像。模型会将所有图像作为参考来生成आउटपुट。
`low` 和 `medium` 速度更快、价格更低,适合草稿和批量编辑。`high` 会运行全保真模型,最适合最终稿。定价会同时随质量和分辨率变化。
通过प्रॉम्प्ट中的自然语言指令,可以进行风格迁移、背景修改、对象修改、构图调整和创意重释。
会。你可以要求模型保留特定元素(例如产品形状或人物),同时修改背景或风格等其他方面。