现已可用最后更新:2026-09-03

GPT Image 2

OpenAI 的图像模型,适合在 Banana Pie 中处理依赖参考图的编辑、产品场景、字体排版测试和高分辨率图像工作。

GPT Image 2 示例输出

关键信息

开发者OpenAI
官方模型 IDgpt-image-2
发布日期2026-04-21
在 Banana Pie 上的可用状态现已可用
所需积分20 积分 / 张
参考图最多 8 张

当图像需要准确遵循参考、材质、文字或布局,而不只是看起来合理时,GPT Image 2 很适合。它不是最便宜的起步选择,所以更适合用于提示词约束足够多、值得进行更细致图像生成的任务。

优势

  • 依赖参考图的编辑

    Banana Pie 标明支持最多 8 张参考图,证据中包括六参考对象绑定、身份与服装迁移,以及在场景变化中保留产品。

  • 产品与光照控制

    测试套件包含产品材质与光照、一致的场景重新布光,以及产品保留,这些都是电商和营销活动图像中的实用检查项。

  • 文字与布局场景

    证据包括多语言字体排版、结构化 UI 图形和图中文字替换,因此当文字和布局属于图像需求的一部分时,它是一个合理选择。

  • 高分辨率输出

    GPT Image 2 标明最高分辨率为 4K,当构图值得进一步打磨时,可用于高分辨率图像输出。

局限

  • 起步成本更高

    最低 20 积分意味着 GPT Image 2 更适合严肃草稿和最终稿,而不是一次性的大量提示词探索。

  • 没有比较胜出声明

    证据运行是按 auto:only-run 规则选择的,因此这些说明展示的是测试覆盖范围,而不是与其他模型的并排排名。

  • 仅图像工作流

    证据只覆盖图像场景,运行事实也将其列为图像模型。

  • 参考图数量有限

    参考图容量最高为 8,因此需要更大批量身份一致性或目录条件控制的项目,可能需要不同的工作流。

适合选择 GPT Image 2 的情况

  • 你需要一个能基于多张参考图工作的图像模型。
  • 你在制作产品、营销活动或目录视觉图,并且材质、光照和保留细节很重要。
  • 你的提示词包含字体排版、类似 UI 的结构,或图中文字编辑。
  • 你想为静态图像获得最高 4K 的输出。

适合考虑其他模型的情况

  • 你只是探索粗略概念,并希望用最低积分成本测试大量提示词。
  • 你需要生成视频,而不是静态图像。
  • 你的图像依赖超过 8 个参考输入。
  • 你需要证明它优于其他模型;现有证据并不能支持这一点。

根据具体任务,以下模型可能更合适: Nano Banana 2 · Seedream 5 Pro

实测结果

所有模型均使用同一套固定提示词,每个模型展示一项已发布的结果,并提供可查看的提示词和设置。

产品材质与光照

提示词

A premium product photograph of exactly one transparent rectangular perfume bottle, half filled with amber liquid, standing upright on wet black stone. Light comes from the upper left, creating coherent refraction, a contact shadow, and one subtle reflection. No text, logo, plants, or extra objects.

设置

  • Resolution: 1K
  • Aspect Ratio: 1:1
多语言排版

提示词

Design a clean cream-colored vertical event poster. Show exactly four centered text lines and no other text: "MOONLIGHT MARKET", "月光市集", "18 OCT", and "RIVER HALL". Render "月光市集" in red and all other lines in black.

设置

  • Resolution: 1K
  • Aspect Ratio: 2:3
人体结构与接触关系

提示词

A photorealistic close-up of an adult potter shaping a clay bowl on a spinning wheel. Both hands are fully visible, each with five natural fingers touching the clay. No other people or hands. Soft window light.

设置

  • Resolution: 1K
  • Aspect Ratio: 3:4
数量识别与属性绑定

提示词

On a matte gray table, exactly three red wooden cubes form a row on the left, exactly two blue glass spheres sit on the right, and one yellow ceramic mug stands centered behind them. The mug handle points right. No other objects or text.

设置

  • Resolution: 1K
  • Aspect Ratio: 4:3
广告构图

提示词

Create a vertical social ad for a fictional running shoe named AERO. Keep the top 15 percent empty. Directly below it, place the headline "RUN LIGHT" in the upper-left. Show exactly one silver shoe in the lower-right, an orange trail curving from the bottom-left, and a round badge reading "42 KM". No other shoes or text.

设置

  • Resolution: 1K
  • Aspect Ratio: 9:16
结构化 UI 图形

提示词

Create a clean horizontal pricing comparison graphic titled "CHOOSE YOUR PLAN". Use exactly three equal columns labeled "STARTER", "PRO", and "TEAM". Under each column, show exactly three aligned rows labeled "PROJECTS", "STORAGE", and "SUPPORT", followed by one blue button labeled "SELECT". White background, dark navy text, no additional columns or text.

设置

  • Resolution: 1K
  • Aspect Ratio: 16:9
六参考图对象绑定与编辑

提示词

Use Reference 1 as the base tabletop scene. Replace the red mug with the exact dark navy-blue mug from Reference 2 in the same position and orientation. Replace the folded white towel with the exact green-and-white striped napkin from Reference 3 in the same folded area. Replace the green pear with exactly one orange from Reference 4 in the same position. Replace the blue notebook with the exact mustard-yellow hardcover notebook from Reference 5 in the same position. Add exactly the two lemons from Reference 6 to the right of the mug. Preserve the wooden table, camera, framing, wood grain, lighting, shadows, and all other spatial relationships from Reference 1. Do not copy the white product backgrounds from References 2–6, duplicate any asset, or add other objects.

设置

  • Resolution: 1K
  • Aspect Ratio: 4:3
图中文字替换

提示词

Replace only the sign text with exactly "NIGHT OWL". Preserve the original font style, spacing, perspective, sign material, lighting, and everything else.

设置

  • Resolution: 1K
  • Aspect Ratio: 3:2
多参考图身份与服装迁移

提示词

Use the portrait in Reference 1 for the person and the isolated jacket in Reference 2 for the garment. Dress the person from Reference 1 in the exact jacket shown in Reference 2. Preserve the person's identity, face, expression, skin, hair, hands, pose, body proportions, background, framing, and lighting from Reference 1. Preserve the jacket's material, color, collar, buttons, pockets, and sleeve patch from Reference 2. Do not copy the ghost mannequin or white product background.

设置

  • Resolution: 1K
  • Aspect Ratio: 3:4
连贯场景重布光

提示词

Change the lighting to warm golden-hour sunlight entering from the left window. Do not move, add, remove, or redesign any object. Update highlights and shadows coherently.

设置

  • Resolution: 1K
  • Aspect Ratio: 16:9
跨比例扩图

提示词

Expand the canvas to a 16:9 landscape by naturally continuing the beach, ocean, and sky on both sides. Keep the complete original image centered without cropping, stretching, letterboxing, or modifying it.

设置

  • Resolution: 1K
  • Aspect Ratio: 16:9
场景变更中的产品保真

提示词

Place the sneaker on a wet outdoor basketball court at dusk. Preserve the exact sneaker shape, black geometric side mark, materials, stitching, sole geometry, and camera angle. Add physically coherent contact, reflections, and dusk lighting.

设置

  • Resolution: 1K
  • Aspect Ratio: 1:1

积分与价格

GPT Image 2 在 Banana Pie 中起步为 20 积分。实际使用时,应把它视为适合有明确目标的图像尝试:当参考图、文字、布局或细节保留重要到足以支撑这笔开销时再使用。

常见问题

GPT Image 2 可以在 Banana Pie 中使用吗?

可以。GPT Image 2 已在 Banana Pie 中列为可用于图像生成。

我可以在 GPT Image 2 中使用参考图吗?

当你需要在一张图像中组合或编辑多张参考图时,可以使用它。Banana Pie 标明支持最多 8 张参考图,测试集包含六参考对象绑定、身份与服装迁移,以及在场景变化中保留产品。

GPT Image 2 可以生成什么分辨率?

Banana Pie 标明 GPT Image 2 的最高分辨率为 4K。

它测试过哪些类型的图像任务?

已发布的测试精选覆盖产品光照、多语言字体排版、结构化 UI 图形、文字替换、重新布光、扩图和基于参考的编辑。它们展示了用于评估该模型的提示词类型,并不保证每一张图像都能达到相同结果。

什么时候不该在它上面花积分?

如果你需要成本很低的草稿、把快速迭代放在首位,或需要视频生成,请选择其他模型。这里的 GPT Image 2 定位是一个谨慎处理图像的模型,适合依赖参考图和对细节敏感的工作。

相关对比

测试方法与披露

结论来自一套固定提示词测试,以及 GPT Image 2 已发布的精选运行结果,覆盖产品材质与光照、多语言字体排版、人体结构与接触关系、计数与属性绑定、广告构图、结构化 UI 图形、基于参考的编辑、文字替换、重新布光、扩图,以及产品保留。

Banana Pie 在同一个创作台中提供此模型及其他模型的付费使用权限。我们的结论来自与用户相同流程下的测试。