GPT Image 2.5 提示词指南
选择模型,编写有效的提示词,并在多次编辑中保留细节。
概览
先明确您需要的图像,再描述主体、构图、风格和约束条件。编辑图像时,请明确哪些内容应当更改,哪些必须保持不变。每次只调整一个方面,并检查结果。
GPT Image 2.5 提供两种模型选择。GPT Image 2.5 Flare 是针对速度优化的小型模型,图像质量与 GPT Image 2 相当。GPT Image 2.5 Sunburst 是针对质量优化的基础模型,图像质量高于 GPT Image 2。两种模型在精确编辑和主体保留方面均有改进。
有关 API 设置和请求示例,请参阅图像生成指南。
选择模型
对于新的工作流程,如果优先考虑速度,请先使用 GPT Image 2.5 Flare;如果优先考虑满足严格的质量要求,请先使用 GPT Image 2.5 Sunburst。输出满足您的要求后,再寻找降低延迟的机会。
从现有图像模型迁移时,请以您当前的图像质量为基准。两种模型均支持图像生成、编辑和透明背景。
| 您当前的工作流程 | 建议首先测试 |
|---|---|
| 现有且经过验证的 GPT Image 2 工作流程已满足您的质量要求 | GPT Image 2.5 Flare。检查能否在降低延迟的同时保持可接受的质量。 |
| GPT Image 2 无法满足您质量要求的复杂用例 | GPT Image 2.5 Sunburst。首先确认它能达到您需要的质量。 |
如果 GPT Image 2.5 Sunburst 满足您的质量要求,请使用相同的提示词和输入测试 GPT Image 2.5 Flare。如果 GPT Image 2.5 Flare 同样满足这些要求,并且能降低延迟,就切换到该模型。如果您的工作流程需要 GPT Image 2.5 Sunburst 的质量优势,请继续使用它。
请使用您自己的工作负载测量响应时间和质量。结果取决于您的提示词、参考图像、输出尺寸和质量设置;某个工作负载的速度有所提升,并不意味着另一个工作负载也能获得固定幅度的提升。
模型参数
请将 API 参数与提示词分开设置。
| 参数 | GPT Image 2.5 设置 |
|---|---|
model | gpt-image-2.5-flare(小型模型)或 gpt-image-2.5-sunburst(基础模型) |
quality | auto(默认)、low、medium、high、xhigh 或 max |
size | auto 或自定义分辨率。常用尺寸:1024x1024(正方形)、1536x1024(横向)、1024x1536(纵向)、2048x2048(2K 正方形)、2048x1152(2K 横向)、3840x2160(4K 横向)和 2160x3840(4K 纵向)。 |
background | auto、opaque 或 transparent |
如需自定义分辨率,请使用 WIDTHxHEIGHT,并遵守以下约束条件:
- 每条边的长度不得超过 3,840 像素。
- 两条边的像素数都必须是 16 的倍数。
- 长边与短边的比例不得超过 3:1。
- 总像素数必须介于 655,360 和 8,294,400 之间。
总像素数超过 3,686,400(2560x1440)的输出属于实验性功能。
在调整 quality 之前,请按照上述工作流程选择模型。首次比较时,如果两种模型都支持您明确选定的质量设置,请保持该设置不变,同时保持提示词、参考图像和输出尺寸不变。相同的质量标签并不意味着不同模型的图像质量或响应时间相同。
如果输出未达要求,请测试更高的质量设置。输出满足您的要求后,请测试较低的设置,看看能否在降低延迟的同时保持可接受的质量。只有当 xhigh 或 max 能在您的延迟预算内改善尚未达标的质量表现时,才使用它们。更高的设置并不能保证每个提示词都能获得更好的结果。
对于透明素材,请明确指定 background="transparent",并使用 PNG 或 WebP。检查解码后图像的 alpha 通道,包括头发、玻璃、阴影和物体边缘。output_compression 仅适用于 JPEG 或 WebP 输出,不适用于 PNG。
迁移现有工作流程
- 保存基准。 收集生产环境中具有代表性的提示词和参考图像,涵盖复杂编辑、精确文本、人脸、产品几何形状和透明素材。记录当前的模型、请求设置和结果。
- 选择首个候选模型。 如果 GPT Image 2 已满足您的质量要求,请先使用 GPT Image 2.5 Flare,测试能否降低延迟。如果 GPT Image 2 在复杂用例中未达要求,请先使用 GPT Image 2.5 Sunburst,并首先确认它能满足您的质量要求。首次比较时,保持提示词、参考图像、尺寸和输出格式不变。
- 全面检查结果。 比较指令遵循情况、人物身份特征和产品的保留情况、文本准确性、非预期改动以及透明效果。重复发送请求以测量一致性。对于编辑工作流程,既要测试完整的编辑序列,也要测试各个步骤。
- 质量达标后,再测试能否降低延迟。 如果您先使用了 GPT Image 2.5 Sunburst,且它满足您的质量要求,请按照相同要求评估 GPT Image 2.5 Flare。只有在质量仍然可接受且延迟有所降低时才切换;否则,请继续使用 GPT Image 2.5 Sunburst。
- 每次只调整一项设置。 在改写提示词之前,先比较不同质量级别。测量常规响应和慢响应的耗时、失败情况、重试情况,以及每张验收合格图像的成本。请确认当前定价,不要假定速度更快的模型成本更低。
- 按工作流程逐步上线。 已发布的模型达到您的验收标准后,先将少量流量切换过去,监测相同的指标,再逐步扩大流量。在旧模型仍受支持期间,保留它以便回滚。
从 GPT Image 1 或 1.5 迁移时,请在参考资料选项卡中查看参数差异和停用日期。请测试候选模型支持的请求设置,不要原样照搬旧设置。从 GPT Image 2 迁移时,请在比较中保留现有的分辨率和透明度要求。
反复编辑仍可能改变您希望保留的细节。请重申这些约束条件,并检查每次结果。如果某个区域必须保持像素完全一致,请将确认合格的编辑部分合成到原始图像中,而不要仅依赖提示词。
提示词基础
- 明确预期结果。 说明主体和预期用途,例如产品照片、广告或示意图。指定构图、宽高比和重要的位置约束。对于复杂请求,请将提示词按场景、主体、细节和约束条件分节组织,并为各节添加标签。
- 选择易于维护的格式。 简短提示词、描述性段落、类 JSON 结构、指令和标签都能表达相同的意图。请选择最便于阅读和更新需求的格式,而不要依赖特殊语法。
- 描述可见细节。 说明材质、光照、颜色和视觉表现媒介。如果目标是照片效果,请明确要求“照片级写实”或“真实照片”,并描述取景和质感。相机参数可用来提示外观效果,但不能保证精确的物理模拟。对于宽幅、电影感、弱光、雨天或霓虹场景,请明确描述尺度、氛围和色彩,不要只依赖描述情绪的词语。
- 明确人物与动作。 描述身体的取景范围、相对大小、视线方向,以及人物与物体的互动。“全身可见,包括双脚”“低头看着摊开的书”或“双手自然握住车把”等指令,可以更清楚地表达所需的姿势和动作。
- 指定准确文本。 将所需文字放在引号中,并描述其位置和字体排版。必要时,逐字母拼出不常见的单词或品牌名称。要求不要添加额外文字,然后检查输出中的拼写和文字清晰度。对于小字号文本、密集信息或多种字体,请比较中、高两种质量设置的效果。
- 区分改动与约束。 编辑时,请说明“仅更改 X”,并列出需要保留的细节,例如人物身份特征、几何形状、布局、光照或标签。说明需要排除的内容,例如不需要的文本、徽标或水印。对于精确的局部编辑,还应明确需要保持不变的饱和度、对比度、箭头、相机角度和周围物体。
- 明确参考图像的用途。 用编号和用途标识每个输入:主体、风格、服装或背景。说明这些输入应如何组合,以及哪些元素应移动到哪里。
- 有针对性地迭代。 将上一次输出作为下一次编辑的输入,每次请求一项改动,并重申需要保留的细节。“与之前风格相同”等表述可以延续上下文,但如果结果出现偏差,请重申关键约束条件。在添加更多指令之前,先比较结果。
以下示例分别演示不同的技巧。您可以将其中的提示词作为起点,根据自己的图像和需求进行调整。
生成图像
控制风格与光照
从主体、取景、光线和质感几个方面描述照片。本示例指定了抓拍式构图,并明确排除过度修图。
生成设置:size="1024x1536"、quality="medium"。
Create a photorealistic candid photograph of an elderly sailor standing on a small fishing boat.
He has weathered skin with visible wrinkles, pores, and sun texture, and a few faded traditional sailor tattoos on his arms.
He is calmly adjusting a net while his dog sits nearby on the deck. Shot like a 35mm film photograph, medium close-up at eye level, using a 50mm lens.
Soft coastal daylight, shallow depth of field, subtle film grain, natural color balance.
The image should feel honest and unposed, with real skin texture, worn materials, and everyday detail. No glamorization, no heavy retouching.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
用图像解释流程
明确要呈现的流程、目标受众以及图像应传达的信息。对于示意图和信息图,除了检查外观,还要核实标注和图中各项关系是否符合事实。
生成设置:size="1024x1536"、quality="medium"。
Create a detailed Infographic of the functioning and flow of an automatic coffee machine like a Jura.
From bean basket, to grinding, to scale, water tank, boiler, etc.
I'd like to understand technically and visually the flow.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
准确呈现指定文字
用引号括起所需文案,并告诉模型这些文字应出现几次。明确目标受众和视觉表现方式,不要添加无关指令。
生成设置:size="1024x1536"、quality="medium"。
Give me a cool in culture ad / fashion shot for a brand called Thread.
It's a hip young street brand. The ad shows a group of friends hanging out together with the tagline "Yours to Create."
Make it feel like a polished campaign image for a youth streetwear audience: stylish, contemporary, energetic, and tasteful.
Use clean composition, strong color direction, natural poses, and premium fashion photography cues.
Render the tagline exactly once, clearly and legibly, integrated into the ad layout.
No extra text, no watermarks, no unrelated logos.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
设计可重复使用的标志
描述品牌,以及构成标志的主要形状。明确要求构图清晰,并在不同尺寸下都易于辨认。使用 n 请求生成多个变体。
生成设置:size="1024x1536"、quality="medium"、background="transparent"、output_format="png"、n=1。
Create an original, non-infringing logo for a company called Field & Flour, a local bakery.
The logo should feel warm, simple, and timeless. Use clean, vector-like shapes, a strong silhouette, and balanced negative space.
Favor simplicity over detail so it reads clearly at small and large sizes. Flat design, minimal strokes, no gradients unless essential.
Fully transparent background. Deliver a single centered logo with generous padding, clean alpha edges, and no solid backdrop, scenery, checkerboard, or watermark.
每行对比两个模型各自生成的一个变体。
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
运用历史与现实背景
指定地点和日期,确立历史背景。模型可以推断相关背景细节,但您仍需检查服装、场景布置和周围环境是否符合史实。
生成设置:size="1024x1536"、quality="medium"。
Create a realistic outdoor crowd scene in Bethel, New York on August 16, 1969.
Photorealistic, period-accurate clothing, staging, and environment.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
将故事转化为连环漫画
根据故事生成漫画时,将叙事拆解为一连串清晰的画面,每格呈现一个情节节点。描述应具体并以动作为重点,让模型能够将故事转化为易于理解、节奏得当的分格画面。
生成设置:size="1024x1536"、quality="medium"。
Create a short vertical comic-style reel with 4 panels.
Panel 1: The owner leaves through the front door. The pet is framed in the window behind them, small against the glass, eyes wide, paws pressed high, the house suddenly quiet.
Panel 2: The door clicks shut. Silence breaks. The pet slowly turns toward the empty house, posture shifting, eyes sharp with possibility.
Panel 3: The house transformed. The pet sprawls across the couch like it owns the place, crumbs nearby, sunlight cutting across the room like a spotlight.
Panel 4: The door opens. The pet is seated perfectly by the entrance, alert and composed, as if nothing happened.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
创建界面预览
像描述已有产品一样描述您想要的产品,生成的界面预览效果最好。重点说明布局、层级、间距和实际界面元素,避免使用概念艺术类表述,让结果看起来像已经发布、可供使用的界面,而不是设计草图。
生成设置:size="1024x1536"、quality="medium"。
Create a realistic mobile app UI mockup for a local farmers market.
Show today’s market with a simple header, a short list of vendors with small photos and categories, a small “Today’s specials” section, and basic information for location and hours.
Design it to be practical, and easy to use. White background, subtle natural accent colors, clear typography, and minimal decoration.
It should look like a real, well-designed, beautiful app for a small local market.
Place the UI mockup in an iPhone frame.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
创建科学与教育类视觉素材
科学与教育类视觉素材非常适合用于生物、化学、课堂讲解、扁平化科学图标体系、图示和学习材料。编写提示时,请像撰写教学设计简报一样,明确受众、教学目标、视觉形式、所需标签和科学性要求。为获得最佳效果,请要求采用简洁、扁平化的视觉体系,保持图标风格统一、箭头清晰、标签易读,并留出足够空白,让学生能快速浏览并理解概念。
如果准确性很重要,请明确列出所需组成部分,并说明不应包含的内容。对于标签密集的图像、图示,或将用于幻灯片或课程材料的素材,请使用 quality="high"。
生成设置:size="1536x1024"、quality="high"。
Create a simple biology diagram titled "Cellular Respiration at a Glance" for high school students.
Show how glucose turns into energy inside a cell. Include glycolysis, the Krebs cycle, and the electron transport chain.
Use arrows to connect the steps, and label the main molecules: glucose, pyruvate, ATP, NADH, FADH2, CO2, O2, and H2O.
Make it look like a clean classroom handout or slide, with a white background, simple icons, clear labels, and easy-to-read text.
Avoid tiny text, extra decoration, or anything that makes the diagram hard to understand.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
制作幻灯片、图示和图表
制作办公类视觉素材时,将提示写成成品规格说明,而非插画需求,效果更好。请明确具体交付内容(幻灯片、工作流程图、图表或页面图像),定义画布和信息层级,提供实际使用的文字或数据,并描述视觉语言。这类提示应包含实用的约束条件:字体排版清晰易读、间距精心安排、无杂乱装饰,也不采用千篇一律的图库照片风格。
对于幻灯片、图表和以图示为主的素材,请直接在提示中提供数字和标签。演示文稿类输出应使用横向尺寸;如果图像包含小字号文字、图例、坐标轴或脚注,请使用 quality="high"。
以下示例中的市场数据和引用来源均为虚构,仅供设计使用。使用该幻灯片前,请将其替换为经过核实的数据。
生成设置:size="1536x864"、quality="high"。
Create one pitch-deck slide titled **"Market Opportunity"** that feels like a real Series A fundraising slide from a YC-backed startup.
Use a clean white background, modern sans-serif typography like Inter, and a crisp, minimal layout. The slide should include:
* A TAM/SAM/SOM concentric-circle diagram in muted blues and grays
* Specific, believable market sizing numbers:
* **TAM:** $42B
* **SAM:** $8.7B
* **SOM:** $340M
* A clean bar chart below showing market growth from **2021 to 2026**, with a subtle upward trend
* Small footnotes: **"AGI Research, 2024"** and **"Internal analysis"**
* A company logo placeholder in the bottom-right corner
The design should look like it belongs in a deck that actually raised money: highly readable text, clear data hierarchy, polished spacing, and professional startup-style visual language.
Avoid clip art, stock photography, gradients, shadows, decorative elements, or anything that feels generic or overdesigned.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
编辑图像
请将引用的输入图像传入 client.images.edit。对于需要蒙版的局部编辑,请参阅使用蒙版编辑图像。
翻译文字并保留布局
请使用各模型在用图像解释流程中生成的咖啡机示意图作为输入。要求替换图中的文字,同时保持设计不变,然后检查译文以及是否有文字仍保留原语言。
编辑设置:size="1024x1536"、quality="high"。
Translate the text in the infographic to Spanish. Do not change any other aspect of the image.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
迁移视觉风格
明确参考图像的具体用途:参考其配色、纹理或视觉媒介。单独描述新的主体。使用下方的像素画作为输入图像。
编辑设置:size="1024x1536"、quality="medium"。
Use the same style from the input image and generate a man riding a motorcycle on a white background.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
保持人物特征不变并更换服装
使用下方的人物照片和三张服装参考图作为输入。说明人物的哪些特征必须保持不变,并限定只更换服装。这种方法也适用于需要保持产品或物体可辨识特征的编辑任务。
编辑设置:size="1024x1536"、quality="medium"。
Edit the image to dress the woman using the provided clothing images. Do not change her face, facial features, skin tone, body shape, pose, or identity in any way. Preserve her exact likeness, expression, hairstyle, and proportions. Replace only the clothing, fitting the garments naturally to her existing pose and body geometry with realistic fabric behavior. Match lighting, shadows, and color temperature to the original photo so the outfit integrates photorealistically, without looking pasted on. Do not change the background, camera angle, framing, or image quality, and do not add accessories, text, logos, or watermarks.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
组合参考图像
将场景照片作为图像 1 传入,将狗的照片作为图像 2 传入。明确要移动哪个元素、将其移到哪里,以及哪些内容必须保持不变。
编辑设置:size="1024x1536"、quality="medium"。
Place the dog from the second image into the setting of image 1, right next to the woman, use the same style of lighting, composition and background. Do not change anything else.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
制作透明背景的产品抠图
在提示中要求单独提取主体,同时在 API 中设置 background="transparent"。使用 PNG 或 WebP,保留返回的 alpha 通道;使用 PNG 时,请省略 output_compression。绘制的棋盘格背景并不是真正的透明背景。在后续编辑中,请重复说明保留透明背景的要求。使用下方的产品照片作为输入。
编辑设置:size="1024x1536"、quality="medium"、background="transparent"、output_format="png"。
Extract the product from the input image and isolate it on a fully transparent background.
Output: centered product, crisp silhouette, no halos/fringing.
Preserve product geometry and label legibility exactly.
Add only light polishing. Do not add a solid backdrop, checkerboard, scenery, or shadow.
Do not restyle the product; remove the background and preserve clean alpha transparency.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
将绘图转换为逼真图像
从草图到渲染图的工作流非常适合将粗略绘图转换为照片级逼真的概念图,同时保留原始意图。请像编写规格说明一样编写提示:保留布局和透视,再通过指定合理的材质、光照和环境来 增强真实感 。加入“不要添加新元素或文字”的要求,以避免模型进行创造性的重新诠释。
编辑设置:size="1024x1536"、quality="medium"。
Turn this drawing into a photorealistic image.
Preserve the exact layout, proportions, and perspective.
Choose realistic materials and lighting consistent with the sketch intent.
Do not add new elements or text.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
移除物体
明确指出要移除的物体,并保留其周围的一切内容。保持人物、姿势、光照和构图不变,确保编辑仅限于局部。
编辑设置:size="1024x1536"、quality="medium"。
Remove the flower from man's hand. Do not change anything else.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
将人物放入场景
将人物放入新场景,同时保留其身份特征。明确指定自然的光照、真实可信的细节、身体取景范围、视线方向以及与场景的互动。说明哪些面部特征和比例必须保持不变。使用 gpt-image-2 时,请省略 input_fidelity;图像输入始终以高保真度处理。
使用博物馆中的女子作为输入图像。
编辑设置:size="1024x1536"、quality="medium"。
Generate a highly realistic action scene where this person is running away from a large, realistic brown bear attacking a campsite. The image should look like a real photograph someone could have taken, not an overly enhanced or cinematic movie-poster image.
She is centered in the image but looking away from the camera, wearing outdoorsy camping attire, with dirt on her face and tears in her clothing. She is clearly afraid but focused on escaping, running away from the bear as it destroys the campsite behind her.
The campsite is in Yosemite National Park, with believable natural details. The time of day is dusk, with natural lighting and realistic colors. Everything should feel grounded, authentic, and unstyled, as if captured in a real moment. Avoid cinematic lighting, dramatic color grading, or stylized composition.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
通过多轮交互完善图像
从一张输出图像开始,检查结果,再将其用作下一次输入。每次后续请求都应限定在较小范围内,以便您看出哪项修改起了作用。
创建初始图像
使用创建透明背景的产品抠图中的洗发水照片作为此广告牌场景的输入。用引号逐字引用标签文字。
编辑设置:size="1024x1536"、quality="medium"。
Create a realistic billboard mockup of the shampoo on a highway scene during sunset.
Billboard text (EXACT, verbatim, no extra characters):
"Fresh and clean"
Typography: bold sans-serif, high contrast, centered, clean kerning.
Ensure text appears once and is perfectly legible.
No watermarks, no logos.
输入图像:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
改变一个条件
将每个模型在上一步输出的广告牌图像传入该模型的下一次编辑请求。这条简短的后续请求会改变天气,同时保留现有场景。
编辑设置:size="1024x1536"、quality="medium"。
Make it look like a winter evening with snowfall.
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
保持角色一致
为包含多幅插图的书籍创建可重复使用的角色参考图,有助于在不同场景、姿势和页面中保持角色外观一致。改变环境和故事情节时,应重复说明角色的标志性细节。
确立角色形象
明确角色的外观、身体比例、服装和整体基调。
生成设置:size="1024x1536"、quality="medium"。
Create a children’s book illustration introducing a main character.
Character:
A young, storybook-style hero inspired by a little forest outlaw,
wearing a simple green hooded tunic, soft brown boots, and a small belt pouch.
The character has a kind expression, gentle eyes, and a brave but warm demeanor.
Carries a small wooden bow used only for helping, never harming.
Theme:
The character protects and rescues small forest animals like squirrels, birds, and rabbits.
Style:
Children’s book illustration, hand-painted watercolor look,
soft outlines, warm earthy colors, whimsical and friendly.
Proportions suitable for picture books (slightly oversized head, expressive face).
Constraints:
- Original character (no copyrighted characters)
- No text
- No watermarks
- Plain forest background to clearly showcase the character
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
延续故事
复用各模型生成的角色图片,并描述一个新场景。再次明确外观约束,确保角色保持一致。
编辑设置:size="1024x1536"、quality="medium"。
Continue the children’s book story using the same character.
Scene:
The same young forest hero is gently helping a frightened squirrel
out of a fallen tree after a winter storm.
The character kneels beside the squirrel, offering reassurance.
Character Consistency:
- Same green hooded tunic
- Same facial features, proportions, and color palette
- Same gentle, heroic personality
Style:
Children’s book watercolor illustration,
soft lighting, snowy forest environment,
warm and comforting mood.
Constraints:
- Do not redesign the character
- No text
- No watermarks
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
更多工作流
更换房间内的家具
直观呈现真实空间中更换家具或装饰后的效果,无需重新创建整个场景。目标是通过精准的局部修改保持真实感:仅替换一个物件,同时保留拍摄角度、光线、阴影和周围环境,让编辑结果看起来像真实照片,而非重新设计的场景。
编辑设置:size="1536x1024"、quality="medium"。
In this room photo, replace ONLY the white chairs with chairs made of wood.
Preserve camera angle, room lighting, floor shadows, and surrounding objects.
Keep all other aspects of the image unchanged.
Photorealistic contact shadows and fabric texture.
输入图片:
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
设计节日贺卡
构思节日贺卡时,请描述场景、情感基调、材质、光线和确切文案。如果希望呈现 3D 立体贺卡或贺卡实拍效果,请明确说明纸张层次、纤维、折痕和柔和的摄影棚灯光。下面的示例采用了带有怀旧气息的泰迪熊场景。
生成设置:size="1024x1536"、quality="medium"。
Create a Christmas holiday card illustration.
Scene:
a cozy Christmas scene with an old teddy bear sitting inside a keepsake box, slightly worn fur, soft stitching repairs, placed near a window with falling snow outside. The scene suggests the child has grown up, but the memories remain.
Mood:
Warm, nostalgic, gentle, emotional.
Style:
Premium holiday card photography, soft cinematic lighting,
realistic textures, shallow depth of field,
tasteful bokeh lights, high print-quality composition.
Constraints:
- Original artwork only
- No trademarks
- No watermarks
- No logos
Include ONLY this card text (verbatim):
"Merry Christmas — some memories never fade."
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
设计收藏类商品
从产品摄影中的材质、包装和印刷清晰度等要素入手,探索商品和包装的设计方案。确保设计原创且不侵权,并比较多种角色或包装方案。
生成设置:size="1024x1536"、quality="medium"。
Create a collectible action figure of a vintage-style toy propeller airplane with rounded wings, a front-mounted spinning propeller, slightly worn paint edges, classic childhood proportions, designed as a nostalgic holiday collectible, in blister packaging.
Concept:
A nostalgic holiday collectible inspired by the simple toy airplanes
children used to play with during winter holidays.
Evokes warmth, imagination, and childhood wonder.
Style:
Premium toy photography, realistic plastic and painted metal textures,
studio lighting, shallow depth of field,
sharp label printing, high-end retail presentation.
Constraints:
- Original design only
- No trademarks
- No watermarks
- No logos
Include ONLY this packaging text (verbatim):
"Christmas Memories Edition"
输出示例:
GPT Image 2.5 Flare
GPT Image 2.5 Sunburst
运行完整示例
此可运行示例仍固定使用 gpt-image-2。请以此为基准,再选择可用模型及其支持的请求设置进行评估。
以下示例会生成四种标志变体,并将产品提取到透明背景上。安装 OpenAI SDK 时,Python 使用 pip install openai,Ruby 使用 gem install openai。请设置 OPENAI_API_KEY,并将产品照片保存为 input_images/shampoo.webp。实际发送请求会产生 API 使用费用。
如需更多提示和完整工作流,请参阅原始笔记本。
检查结果
使用输出前,请对照要求进行检查:
- 所需文字是否准确、清晰可读?图解中的标签和关系是否正确?
- 人物身份特征、产品形状、标签和参考图中的细节是否完整保留?
- 此次编辑是否只修改了您要求修改的内容?
- 如果需要透明效果,文件是否包含 alpha 通道,而不是仅绘制了一个背景?
更改提示或模型时,请使用有代表性的输入比较质量、延迟和成本。当前费用请参阅图像生成定价。
GPT Image 2 参考资料
现有 GPT Image 2 工作流的概览和请求设置。
概览
GPT Image 2 支持图像生成和编辑,包括文字渲染、基于参考图像的编辑以及灵活的输出尺寸。您可以使用本参考资料维护现有集成。提示词指南介绍了构图、文字、参考图像以及在编辑过程中保留细节的通用技巧。其中的图示示例使用 GPT Image 2.5 Flare 和 GPT Image 2.5 Sunburst;不同模型的输出可能有所不同。如需迁移,请参阅该指南中的模型选择和评估工作流程。
模型参数
使用 client.images.generate 生成图像,使用 client.images.edit 编辑图像。有关 API 设置和请求示例,请参阅图像生成指南。
| 参数 | GPT Image 2 |
|---|---|
model | gpt-image-2 |
quality | low、medium、high 或 auto |
size | auto 或支持的分辨率;请参阅尺寸限制 |
input_fidelity | 请省略此参数。输入图像始终以高保真度处理。 |
output_format | png、jpeg 或 webp |
background | 如需输出透明图像,请显式设置为 transparent,并使用 PNG 或 WebP 格式。 |
output_compression | 仅用于 JPEG 或 WebP 输出,不用于 PNG。 |
gpt-image-2 的透明背景功能已开放预览。
有关原始提示、输入和可运行的工作流,请参阅固定版本的 GPT Image 2 笔记本。
GPT Image 1.5 参考资料
现有 GPT Image 1.5 工作流的概览和请求设置。
概览
已弃用的模型。 gpt-image-1.5 计划于 2026 年 12 月 1 日
停止服务。请参阅弃用
通知,并在迁移前
使用 gpt-image-2 验证现有工作流。
GPT Image 1.5 支持图像生成和编辑,包括文字渲染、照片级真实感图像以及基于参考图像的编辑。您可以使用本参考资料维护现有集成。提示词指南介绍了构图、文字、参考图像以及在编辑时保留细节的通用技巧。请使用您的模型和输入测试这些技巧;不同模型的输出可能有所不同。
模型参数
使用 client.images.generate 生成图像,使用 client.images.edit 编辑图像。有关 API 设置和请求示例,请参阅图像生成指南。
| 参数 | GPT Image 1.5 |
|---|---|
model | gpt-image-1.5 |
quality | low、medium、high 或 auto |
size | 1024x1024、1024x1536、1536x1024 或 auto |
output_format | png、jpeg 或 webp |
output_compression | 0 到 100,仅适用于 JPEG 或 WebP 输出 |
background | 如需透明背景输出,请显式设置为 transparent,并使用 PNG 或 WebP 格式 |
input_fidelity | low 或 high;high 保留输入细节,而 quality 控制输出生成。迁移到 GPT Image 2 时,请省略此参数,因为该模型始终采用高输入保真度。 |
GPT Image 1 参考资料
现有 GPT Image 1 工作流的概览和请求设置。
概览
此模型已弃用。 gpt-image-1 计划于 2026 年
10 月 23 日停止服务。请参阅弃用
公告,并在迁移前
使用 gpt-image-2 验证现有工作流。
GPT Image 1 支持使用参考图像和蒙版生成与编辑图像。您可以使用本参考资料维护现有集成。有关描述场景、保留细节和优化编辑效果等通用技巧,请参阅提示词指南。
模型参数
使用 client.images.generate 生成图像,使用 client.images.edit 编辑图像。有关 API 设置和请求示例,请参阅图像生成指南。
| 参数 | GPT Image 1 |
|---|---|
model | gpt-image-1 |
quality | low、medium、high 或 auto |
size | 1024x1024、1024x1536、1536x1024 或 auto |
output_format | png、jpeg 或 webp |
output_compression | 0 至 100,仅适用于 JPEG 或 WebP 输出 |
background | 如需透明输出,请显式设置为 transparent,并使用 PNG 或 WebP 格式 |
input_fidelity | low 或 high;high 用于保留输入细节,quality 则控制输出生成。高输入保真度会消耗更多图像输入 Token。迁移到 GPT Image 2 时,请省略此参数,因为该模型始终使用高输入保真度。 |