使用 GPT Image 2 和文本指令转换并编辑现有图像。支持最多 16 张entrada图像,可进行精准的风格迁移、编辑和图像变换。
About this model
GPT Image 2 imagen a imagen可使用自然语言指令转换和编辑现有图像,支持最多 16 张entrada图像,适用于精准编辑、风格迁移和创意变换。它能够遵循详细指令修改构图、风格和内容,同时保留重要元素;支持选择质量(low / medium / high)以及 1K / 2K / 4K 分辨率。
Cost analysis
| Provider | Cost | Notes |
|---|---|---|
| muapiapp | $0.025 – $0.150 per image | 只需为实际使用付费,与输入图像数量无关。低质量:$0.025 / $0.040 / $0.075(1K / 2K / 4K)。中质量:$0.030 / $0.045 / $0.090。高质量:$0.060 / $0.090 / $0.150。 |
| Fal.ai | $0.040 – $0.110 per image | 仅支持低质量和中等质量。低/中等:$0.040 / $0.060 / $0.110(1K / 2K / 4K)。 |
| Replicate | No disponible | Replicate 不提供 GPT Image 2。 |
只需为实际使用付费,与输入图像数量无关。低质量:$0.025 / $0.040 / $0.075(1K / 2K / 4K)。中质量:$0.030 / $0.045 / $0.090。高质量:$0.060 / $0.090 / $0.150。
仅支持低质量和中等质量。低/中等:$0.040 / $0.060 / $0.110(1K / 2K / 4K)。
Replicate 不提供 GPT Image 2。
** Competitor pricing is estimated based on similar model architectures and usage tiers.
Configuration schema
| Parameter | Type | Description | Default |
|---|---|---|---|
| Prompt | string | 描述所需变换的文本指令。最多 20,000 个字符。 | Preserve the original subject identity, facial structure, hairstyle, pose, body proportions, camera framing, and overall composition from the source image. Transform the environment into a flooded underwater railway terminal with giant glass ceilings showing whales and ocean water above. Add realistic cinematic water reflections, soft underwater caustic lighting, floating atmospheric particles, subtle wetness on clothing and skin, and highly detailed environmental storytelling elements such as abandoned luggage, vines, cracked marble, and cinematic fog depth. Maintain natural realism and believable anatomy while enhancing texture fidelity, lighting realism, color harmony, and cinematic atmosphere. Keep the face sharp and recognizable with authentic skin detail and emotionally grounded expression. Avoid over-stylization, distorted anatomy, plastic skin, blurry textures, extra limbs, cartoon aesthetics, oversaturated colors, low-detail backgrounds, or artificial-looking lighting. |
| Imagen URL | array | Upload or provide input images to transform. Up to 16 images supported. | https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/gpt-image-2-image-to-image-in.jpg |
| Relación de aspecto | Enum (6 options) | salidaimagenrelación de aspecto。注意:relación de aspecto为 'auto'(或未指定)的imagen只能转换为 1K;1:1 不能转换为 4K,否则任务将创建失败。 | auto |
| Resolución | Enum (3 options) | imagenresolución。注意:relación de aspecto为 1:1 的imagen不能转换为 4K;relación de aspecto为 'auto'(或未指定)的imagen只能转换为 1K,否则任务将创建失败。 | 2K |
| Calidad | Enum (3 options) | 生成质量。'low' 和 'medium' 使用更快、更便宜的后端;'high' 使用更高保真的后端。 | high |
描述所需变换的文本指令。最多 20,000 个字符。
Preserve the original subject identity, facial structure, hairstyle, pose, body proportions, camera framing, and overall composition from the source image. Transform the environment into a flooded underwater railway terminal with giant glass ceilings showing whales and ocean water above. Add realistic cinematic water reflections, soft underwater caustic lighting, floating atmospheric particles, subtle wetness on clothing and skin, and highly detailed environmental storytelling elements such as abandoned luggage, vines, cracked marble, and cinematic fog depth. Maintain natural realism and believable anatomy while enhancing texture fidelity, lighting realism, color harmony, and cinematic atmosphere. Keep the face sharp and recognizable with authentic skin detail and emotionally grounded expression. Avoid over-stylization, distorted anatomy, plastic skin, blurry textures, extra limbs, cartoon aesthetics, oversaturated colors, low-detail backgrounds, or artificial-looking lighting.Upload or provide input images to transform. Up to 16 images supported.
https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/gpt-image-2-image-to-image-in.jpgsalidaimagenrelación de aspecto。注意:relación de aspecto为 'auto'(或未指定)的imagen只能转换为 1K;1:1 不能转换为 4K,否则任务将创建失败。
autoimagenresolución。注意:relación de aspecto为 1:1 的imagen不能转换为 4K;relación de aspecto为 'auto'(或未指定)的imagen只能转换为 1K,否则任务将创建失败。
2K生成质量。'low' 和 'medium' 使用更快、更便宜的后端;'high' 使用更高保真的后端。
highDeveloper documentation
上传图像:通过 images_list 字段提供最多 16 张entrada图像。这些图像将作为转换对象。
编写prompt:描述你想要的变换。请具体说明目标风格、修改内容或效果。
选择质量:选择 low 或 medium,以快速完成成本更低的编辑;选择 high,以获得最高保真度的salida。high 为默认值。
选择宽高比和分辨率:选择宽高比(例如 1:1、16:9、9:16)以及分辨率(1K、2K 或 4K)。请注意,1:1 不能以 4K 渲染,而 auto 宽高比固定使用 1K。
提交并查看:点击 Generate。模型会将你的指令应用于entrada图像,并返回转换后的结果。
Frequently asked
你可以在 `images_list` 字段中提供最多 16 张图像。模型会将所有图像作为参考来生成salida。
`low` 和 `medium` 速度更快、价格更低,适合草稿和批量编辑。`high` 会运行全保真模型,最适合最终稿。定价会同时随质量和分辨率变化。
通过prompt中的自然语言指令,可以进行风格迁移、背景修改、对象修改、构图调整和创意重释。
会。你可以要求模型保留特定元素(例如产品形状或人物),同时修改背景或风格等其他方面。