
科學資訊圖表
1:1A detailed scientific infographic of an integrated biohybrid artificial photosynthesis platform for solar-to-fuel conversion, with labeled cross-section diagrams, molecular structures, efficiency charts, and process flow annotations
openai/gpt-image-2OpenAI 的 GPT Image 2 圖片產生 API - 最高 4K,原生推理、多語言文字渲染、參考圖引導編輯。

A detailed scientific infographic of an integrated biohybrid artificial photosynthesis platform for solar-to-fuel conversion, with labeled cross-section diagrams, molecular structures, efficiency charts, and process flow annotations

A photorealistic candid shot of a young man in a light grey Covernat hoodie sitting at station 139 in a premium PC cafe, focused on his laptop screen, soft window light mixing with monitor glow, shallow depth of field

A pixel-perfect recreation of the YouTube homepage UI with a left sidebar showing Home, Shorts, Subscriptions, History, and Explore sections, a top navigation bar with search and profile icon, category filter chips, and an 8-video thumbnail grid with realistic titles, channel names, view counts, and duration stamps
GPT Image 2 是 OpenAI 的旗艦圖片產生模型,於 2026 年 4 月發布。是首批內建推理能力的圖片模型之一 - 模型在產生前會規劃構圖並驗證提示詞的約束條件。支援跨文字系統的多語言文字渲染,包含 CJK、印地語與孟加拉語,非常適合全球化的創意工作。支援最高 4K 輸出,可透過 GPT Image 2 API 存取。
具備精確光線與材質的產品攝影。注重寫實感與精細細節的行銷活動與編輯視覺素材。需要清晰嵌入文字的資訊圖表、UI 模型與海報設計。參考圖引導的編輯 - 上傳最多 4 張圖片來引導風格、色彩或構圖。
所有參數都在執行請求的 input 物件中傳遞。
| Parameter | Required | Description |
|---|---|---|
| prompt | Yes | 圖片的文字描述(1–4000 字元) |
| aspect_ratio | No | 輸出寬高比。預設 1:1。選項:1:1、2:3、3:2、3:4、4:3、4:5、5:4、9:16、16:9、21:9 |
| resolution | No | 輸出解析度。預設 1K。選項:1K、2K、4K(4K 僅限 16:9、9:16 或 21:9) |
| image_urls | No | 最多 4 張參考圖(每張最大 4 MB)用於圖片轉圖片產生 |
GPT Image 2 對物理描述的回應很好。「Matte ceramic vase on a walnut table, soft window light from the left」的效果持續優於「a vase on a table」。
4K 輸出限定為 16:9、9:16 與 21:9。要在橫向或超寬構圖中獲得最大細節,將 resolution: "4K" 與這些寬高比之一搭配使用。
提供與目標風格接近的參考圖 - 模型會用它們來引導色彩與構圖,而非從差異較大的參考中推斷。
十種選項:1:1、2:3、3:2、3:4、4:3、4:5、5:4、9:16、16:9 與 21:9。預設為 1:1。
支援 - 1K、2K 與 4K 皆可用。4K 限定為 16:9、9:16 與 21:9 寬高比。
可以。透過 image_urls 上傳最多 4 張參考圖,用你自己的視覺素材引導產生。
GPT Image 2 是 OpenAI 的旗艦模型,具備內建推理 - 在產生前規劃構圖並驗證提示詞約束條件 - 以及包含 CJK 的強大多語言文字渲染。Nano Banana Pro 是 Google 保真度最高的模型,採用 Gemini 3 Pro,具備 4K 輸出、11 種寬高比與最多 8 張參考圖。追求推理驅動構圖與多語言文字選 GPT Image 2;追求透過 API 進行 4K 輸出與更多參考圖引導編輯選 Nano Banana Pro。
GPT Image 2 是較新的旗艦模型,具備內建推理與改良的多語言文字渲染。GPT Image 1.5 是前一代 - 能力不錯但沒有規劃/驗證步驟。受益於推理的複雜提示詞選 GPT Image 2;不需要推理負擔的簡單產生則適合 GPT Image 1.5。
在 Runbase 上有兩種方式:使用本頁頂部的 Playground,以提示詞與選用的參考圖即時產生;或呼叫 API,向 /api/v1/runs 發送 POST 請求,帶上 model: "openai/gpt-image-2" 與包含 prompt(必填)、aspect_ratio、resolution 及最多 4 個 image_urls 的 input 物件。新帳戶可獲得註冊贈送額度來試用。完整程式碼範例請見 API 分頁。
GPT Image 2 按產生的每張圖片計費,分為三個解析度等級:1K(最便宜)、2K 與 4K(最高)。目前的每張圖片費率請見上方的定價區段。
採按量計費而非免費,但新的 Runbase 帳戶可獲得註冊贈送額度,可用於 GPT Image 2 及其他模型。額度用完後,依定價區段所示的費率按每次產生付費。