关于此模型
GPT Image 2 图生图可使用自然语言指令转换和编辑现有图像,支持最多 16 张输入图像,适用于精准编辑、风格迁移和创意变换。它能够遵循详细指令修改构图、风格和内容,同时保留重要元素;支持选择质量(low / medium / high)以及 1K / 2K / 4K 分辨率。
成本分析
| 提供商 | 费用 | 备注 |
|---|---|---|
| muapiapp | $0.025 – $0.150 每张图像 | 只需为实际使用付费,与输入图像数量无关。低质量:$0.025 / $0.040 / $0.075(1K / 2K / 4K)。中质量:$0.030 / $0.045 / $0.090。高质量:$0.060 / $0.090 / $0.150。 |
| Fal.ai | $0.040 – $0.110 每张图像 | 仅支持低质量和中等质量。低/中等:$0.040 / $0.060 / $0.110(1K / 2K / 4K)。 |
| Replicate | 暂不可用 | Replicate 不提供 GPT Image 2。 |
只需为实际使用付费,与输入图像数量无关。低质量:$0.025 / $0.040 / $0.075(1K / 2K / 4K)。中质量:$0.030 / $0.045 / $0.090。高质量:$0.060 / $0.090 / $0.150。
仅支持低质量和中等质量。低/中等:$0.040 / $0.060 / $0.110(1K / 2K / 4K)。
Replicate 不提供 GPT Image 2。
** 竞品价格根据相似模型架构和使用层级估算。
配置参数
| 参数 | 类型 | 描述 | 默认值 |
|---|---|---|---|
| 提示词 | string | 描述所需变换的文本指令。最多 20,000 个字符。 | Preserve the original subject identity, facial structure, hairstyle, pose, body proportions, camera framing, and overall composition from the source image. Transform the environment into a flooded underwater railway terminal with giant glass ceilings showing whales and ocean water above. Add realistic cinematic water reflections, soft underwater caustic lighting, floating atmospheric particles, subtle wetness on clothing and skin, and highly detailed environmental storytelling elements such as abandoned luggage, vines, cracked marble, and cinematic fog depth. Maintain natural realism and believable anatomy while enhancing texture fidelity, lighting realism, color harmony, and cinematic atmosphere. Keep the face sharp and recognizable with authentic skin detail and emotionally grounded expression. Avoid over-stylization, distorted anatomy, plastic skin, blurry textures, extra limbs, cartoon aesthetics, oversaturated colors, low-detail backgrounds, or artificial-looking lighting. |
| 图像 URL | array | Upload or provide 输入图像s to transform. 最多 16 images supported. | https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/gpt-image-2-image-to-image-in.jpg |
| 画面比例 | 枚举(6 个选项) | 输出图像画面比例。注意:画面比例为 'auto'(或未指定)的图像只能转换为 1K;1:1 不能转换为 4K,否则任务将创建失败。 | auto |
| 分辨率 | 枚举(3 个选项) | 图像分辨率。注意:画面比例为 1:1 的图像不能转换为 4K;画面比例为 'auto'(或未指定)的图像只能转换为 1K,否则任务将创建失败。 | 2K |
| 质量 | 枚举(3 个选项) | 生成质量。'low' 和 'medium' 使用更快、更便宜的后端;'high' 使用更高保真的后端。 | high |
描述所需变换的文本指令。最多 20,000 个字符。
Preserve the original subject identity, facial structure, hairstyle, pose, body proportions, camera framing, and overall composition from the source image. Transform the environment into a flooded underwater railway terminal with giant glass ceilings showing whales and ocean water above. Add realistic cinematic water reflections, soft underwater caustic lighting, floating atmospheric particles, subtle wetness on clothing and skin, and highly detailed environmental storytelling elements such as abandoned luggage, vines, cracked marble, and cinematic fog depth. Maintain natural realism and believable anatomy while enhancing texture fidelity, lighting realism, color harmony, and cinematic atmosphere. Keep the face sharp and recognizable with authentic skin detail and emotionally grounded expression. Avoid over-stylization, distorted anatomy, plastic skin, blurry textures, extra limbs, cartoon aesthetics, oversaturated colors, low-detail backgrounds, or artificial-looking lighting.Upload or provide 输入图像s to transform. 最多 16 images supported.
https://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/gpt-image-2-image-to-image-in.jpg输出图像画面比例。注意:画面比例为 'auto'(或未指定)的图像只能转换为 1K;1:1 不能转换为 4K,否则任务将创建失败。
auto图像分辨率。注意:画面比例为 1:1 的图像不能转换为 4K;画面比例为 'auto'(或未指定)的图像只能转换为 1K,否则任务将创建失败。
2K生成质量。'low' 和 'medium' 使用更快、更便宜的后端;'high' 使用更高保真的后端。
high开发者文档
上传图像:通过 images_list 字段提供最多 16 张输入图像。这些图像将作为转换对象。
编写提示词:描述你想要的变换。请具体说明目标风格、修改内容或效果。
选择质量:选择 low 或 medium,以快速完成成本更低的编辑;选择 high,以获得最高保真度的输出。high 为默认值。
选择宽高比和分辨率:选择宽高比(例如 1:1、16:9、9:16)以及分辨率(1K、2K 或 4K)。请注意,1:1 不能以 4K 渲染,而 auto 宽高比固定使用 1K。
提交并查看:点击 Generate。模型会将你的指令应用于输入图像,并返回转换后的结果。
常见问答
你可以在 `images_list` 字段中提供最多 16 张图像。模型会将所有图像作为参考来生成输出。
`low` 和 `medium` 速度更快、价格更低,适合草稿和批量编辑。`high` 会运行全保真模型,最适合最终稿。定价会同时随质量和分辨率变化。
通过提示词中的自然语言指令,可以进行风格迁移、背景修改、对象修改、构图调整和创意重释。
会。你可以要求模型保留特定元素(例如产品形状或人物),同时修改背景或风格等其他方面。