部落格

GPT Image 2.5 提示詞指南:Sketch、編輯與影片

最後更新於
GPT Image 2.5 提示詞指南:Sketch、編輯與影片

GPT Image 2.5 更像按製作規格執行簡述,而不是按情緒板發揮。OpenAI 自己的 prompting guide 是最好的一手來源:點明結果、描述可見內容、引用精確文字,編輯時還要說明什麼不能變。

本文把這些指引變成可在 ChatGPT Images 2.5 或 API 裡複用的提示詞,再把靜幀交給 PixVerse 當首幀。模型總覽見 什麼是 GPT Image 2.5?OpenAI 新影像模型詳解。API 路由見 GPT Image 2.5 Flare vs Sunburst:API 該選哪個?。

PixVerse 現在包含 GPT Image 2.5 Flare、GPT Image 2.5 Sunburst 和 GPT Image 2。下面的提示詞模式適用於這些 OpenAI 影像模型。可以在 PixVerse 工作區生成 2.5 靜幀,並在那裡繼續圖片轉影片。如果在 ChatGPT 或 OpenAI API 中生成,則下載畫面再上傳。 官方上線起 14 天內,符合條件的海外 PixVerse 會員曾可免積分使用 Flare 和 Sunburst。截至 2026 年 9 月 27 日,不要假定該優惠仍然有效,也不要把任一模型當成免費或無限制。生成前請檢視工作區顯示的積分價格。 可用性依據 GPT Image 2.5 模型頁 與 PixVerse 2026 年 9 月 20 日的評測。

真正管用的簡述

把 OpenAI 的基本原則壓縮成可複用的塊:

Result: [poster / pack shot / UI mock / comic / video first frame] for [use case]. Subject: [who or what, visible details, pose, scale]. Exact text: “[copy in quotes]” — appear once, no extra words. Composition: [framing, placement, negative space, aspect ratio]. Style and light: [medium, lighting, materials, what “photoreal” means here]. Constraints: [do not change X; no watermark; no extra logos]. Output: [size, transparent or opaque, video-ready / print / social].

相對 GPT Image 2 提示詞,關鍵轉變是:把改動與鎖定分開。 “Make it winter” 是不完整的。“Make it a winter evening with snowfall. Keep the billboard copy, product, camera angle, and layout unchanged” 才是 2.5 風格的編輯。

先選介面,再寫提示詞

介面 什麼時候用 對提示詞的含義
ChatGPT Sketch(@Sketch) 佈局畫出來比寫出來更容易 草繪幾何結構;把風格、材質和排除項寫在文字裡
ChatGPT templates 海報、周邊、產品照片、圖示 回答範本問題;不要對抗版式
ChatGPT comments 對現有靜幀做一處區域性改動 點選區域;寫短指令;不要把整場戲再說一遍
API Flare 通量和延遲 把畫質寫明確(medium/high);快速迭代
API Sunburst 成片和硬編輯 與 Flare 用同一條提示詞以便對比;只有失敗時才提高畫質
PixVerse Flare / Sunburst 希望靜幀留在影片工作區 在 Flare 或 Sunburst 上用同一套結構化簡述,然後無需匯出即可圖片轉影片;GPT Image 2 仍在該工作區

TechRadar 的產品照片範本是讓介面來收集簡述的好例子:上傳產品,選 Clean Studio / Editorial / Lifestyle / Dramatic,選場景,然後生成。當範本表達不了某條約束(精確 SKU 標籤、鎖定機位、不要額外文字)時,再用自訂提示詞。

生成提示詞

這些寫法遵循 OpenAI 文件化的技巧。改名詞,保留結構。

寫實產品主視覺(影片首幀)

Create a photorealistic 16:9 product photograph for an image-to-video ad.
Main subject: a matte black wireless earbud case, lid half open, one earbud hovering just above the cradle.
Exact text: none. No logos, no watermarks, no pack copy.
Composition: product centered, large in frame, generous left-side negative space for later captions, tabletop at chest height.
Style and lighting: clean studio, soft key from camera left, faint rim light, shallow depth of field, real plastic and metal texture, no glamor retouching.
Constraints: both earbuds identical; hinge aligned; no extra props.
Output: 16:9, opaque background, still camera, ready to animate with a slow orbit.

給 PixVerse 用的首幀,主體應足夠大、不被遮擋、光線一致。後面的運動提示詞描述運鏡和材質運動,而不是再寫一遍產品。參見 如何把產品照片做成影片廣告。

帶引用文字的海報

Create a concert poster for a live jazz night.
Exact text (verbatim, once each):
Title: "AFTER HOURS"
Subhead: "Friday 10pm — Room 12"
Footer: "All ages"
Typography: bold condensed sans for the title, smaller grotesque for the subhead, generous tracking, high contrast, fully legible.
Scene: a single trumpet in warm spotlight, smoke catching the beam, deep navy background.
Constraints: no extra copy, no sponsor marks, no QR codes, no watermarks.
Output: 2:3 portrait, print-like layout with clear margins.

給字串加引號。每種只要求出現一次。然後在檔案裡核對拼寫。OpenAI 建議小號或密集文字使用 medium 或 high 畫質。

透明素材

在 API 裡,透明度是參數,不是比喻。設定 background="transparent"、output_format="png" 或 webp,並在提示詞裡說清楚:

Extract the kettle from the input photo and isolate it on a fully transparent background.
Centered product, crisp silhouette, no halos or fringing.
Preserve geometry and label lettering exactly.
Do not add a solid backdrop, checkerboard, drop shadow, or scenery.

OpenAI 的示例對 Logo 和包裝摳圖用同一模式。檢查頭髮、玻璃和邊緣的 alpha 通道。畫出來的棋盤格意味著模型把背景畫上去了。

從草圖到成圖

在 ChatGPT 裡用 @Sketch 畫出佈局,然後:

Use this sketch as the layout and perspective reference.
Turn it into a photorealistic modern living room: light oak floor, sage accent wall, daylight from the window shown in the sketch.
Match the sketch's room proportions and furniture positions.
Do not add furniture, people, or text that is not in the sketch.

在 API 裡,把圖紙作為輸入圖片附上,並用同樣的鎖定措辭。OpenAI 的山谷繪圖示例是經典版本:保住佈局,補上合理的材質和光線,不要發明新元素。

編輯提示詞

2.5 的賣點是區域性控制。把編輯寫成 一處改動 + 一份鎖定清單。

改一個物體

In this kitchen photo, replace ONLY the white chairs with light oak chairs.
Preserve camera angle, room lighting, floor shadows, and surrounding objects.
Keep all other aspects of the image unchanged.
Photorealistic contact shadows and wood grain.

保住身份,換衣服

Dress the person using the provided clothing images.
Do not change face, facial features, skin tone, body shape, pose, or identity.
Preserve likeness, expression, hairstyle, and proportions.
Replace only the clothing; fit garments to the existing pose with realistic fabric.
Match lighting and color temperature to the original photo.
Do not change background, camera angle, or framing. No extra accessories, text, or logos.

翻譯版式文字

Translate the infographic text to Spanish.
Do not change any other aspect of the image: icons, colors, arrows, spacing, or diagram structure.

多輪而不漂移

OpenAI 的廣告牌示例節奏是對的:

  1. 用引用文案搭場景:"Fresh and clean"。
  2. 下一輪:Make it look like a winter evening with snowfall. —— 把上一輪輸出當作輸入傳入。
  3. 一旦開始滑動,就重申關鍵鎖定(文案、產品、機位)。

如果某個區域必須畫素級一致,在編輯器裡合成它。即便有 2.5,提示詞也不是無損補丁工具。

在 ChatGPT 裡,這類編輯優先用 comments:在天空上釘一個點,寫 “snowfall, keep the billboard type.” TechRadar 和 The Verge 都發現批註比重寫整場戲更快。

從靜幀到 PixVerse 影片

2.5 靜幀只是社交或廣告工作流的一半。畫面鎖定之後:

  1. 按交付寬高比匯出(16:9、9:16 或 1:1)。避免事後裁切改變構圖。
  2. 在 PixVerse 中,如果靜幀是在這裡製作的,就直接開啟圖片轉影片。只有來自 ChatGPT 或其他工具時才上傳檔案。
  3. 提示詞寫 運動,不要再寫一遍產品:運鏡、材質運動、光線變化。
  4. 第一次測試把時長壓短,動作站得住再延長或改風格。

疊加在耳機盒靜幀上的運動提示詞示例:

Slow 30-degree orbit around the case, lid already open.
The hovering earbud rotates in place with a faint specular glide.
Soft studio light stays consistent. No new props, no on-screen text, no camera shake.

以角色為主的片段,先用 2.5 生成一致的角色設定圖(同一套服裝、同一套面部鎖定,覆蓋多個姿勢),再在 PixVerse 裡做動畫。參見 如何用 AI 建立一致角色。

如果任務並不特別需要 2.5,在 PixVerse 內用 GPT Image 2、Nano Banana 2 或 Seedream 仍然可以省去額外匯出。如果需要 2.5,Flare 和 Sunburst 已在同一工作區,匯出是可選步驟,不再是必經步驟。

畫質、尺寸與成本旋鈕

這些在 API 裡設定在提示詞 之外:

  • model:gpt-image-2.5-flare 或 gpt-image-2.5-sunburst
  • quality:auto、low、medium、high、xhigh、max
  • size:auto 或 WIDTHxHEIGHT
  • background:auto、opaque、transparent

OpenAI 的建議:先選模型,對比模型時固定畫質,只有任務仍失敗時才提高畫質。xhigh 和 max 相對 GPT Image 2 是新增的;它們不是預設的“讓它更好”開關。

所列 token 標價與 GPT Image 2 相同,但 2.5 的 token 數量 不能 用 GPT Image 2 計算器估算。用自己的尺寸測算時間和成本。細節見 image generation pricing。

做動畫之前先檢查檔案

OpenAI 的驗收清單也是很好的影片門檻:

  • 要求的文字是否準確、可讀?
  • 身份、產品形狀和標籤是否完好?
  • 編輯是否只改了你要求的部分?
  • 若需要透明度,是否有真正的 alpha 通道?
  • 主體是否足夠大,能扛住壓縮和運動?

如果靜幀過不了這份清單,就不要在它上面花影片額度。

常見問題

2.5 的提示詞應該更長嗎?

它們應該更有結構,不一定更長。有名稱的段落勝過一長串形容詞。OpenAI 稱短提示詞、帶標籤的塊,以及類似 JSON 的結構都可以用;選你能維護的格式。

能從 API 提示 Sketch 嗎?

不能用 @Sketch。把圖紙作為參考圖傳入,並說明它控制佈局和透視。

這會取代 GPT Image 2 提示詞指南嗎?

不會。GPT Image 2 深度測評:2026 年提示詞指南與實戰案例 仍然適用於 GPT Image 2,包括在 PixVerse 上。當你使用 Images 2.5、Flare 或 Sunburst(包括這些模型在 PixVerse 工作區中執行時),或需要面向 Sketch 與批註的簡述時,再用本文。

相關資源