2K 輸出與遮罩編輯
原生支援最高 2K(2048px)輸出,色彩中性且準確——gpt-image-1.5 的偏暖色調已不復存在。透過遮罩影像對精確區域進行修補或延展,同時未修改的像素將保持完全一致。

curl --request POST "https://gptproto.com/api/v3/openai/gpt-image-2/text-to-image" \
--header "Authorization: Bearer $GPTPROTO_API_KEY" \
--header "Content-Type: application/json" \
--data '{
"prompt": "A tiny origami fox sailing a teacup across a moonlit puddle",
"n": null,
"quality": "auto",
"size": "auto",
"response_format": "url"
}'先從單次樣本成本開始,再選擇測試預算。GPTProto 費率比標價低 20%。
透過單一 GPTProto 金鑰呼叫 OpenAI 的 gpt-image-2,每 100 萬個 token 收費 $6.4/$24——比 OpenAI 的牌價低 20%。使用相同的模型 ID,無需組織驗證,一個餘額即可共用於 200 多個模型。
原生支援最高 2K(2048px)輸出,色彩中性且準確——gpt-image-1.5 的偏暖色調已不復存在。透過遮罩影像對精確區域進行修補或延展,同時未修改的像素將保持完全一致。

單次呼叫最多可傳入 16 張參考圖像,讓一組作品中的主體身分、風格與產品細節保持一致——適用於連續藝術、目錄拍攝與品牌一致的行銷活動。

能以精確的版面配置渲染拉丁文與中日韓文字中的小型標籤、UI 文案和長篇文字——無需另行排版即可用於客戶專案。這是選擇 gpt image 2 api 而非舊款生成器的核心原因之一。

在渲染前,gpt-image-2 會分析版面配置與限制條件——首個具備 O 系列風格規劃能力的 OpenAI 圖像模型。它能提升資訊圖表、多面板版面與包裝等複雜場景的成功率。

gpt-image-2 is OpenAI's image generation and editing model, released April 21, 2026 as the successor to gpt-image-1 (April 2025) and gpt-image-1.5 (December 2025). It is natively multimodal — image generation is part of the core model rather than a diffusion model bolted onto a language model — which is why it follows long, multi-part prompts and renders readable in-image text more reliably than DALL-E-era generators.
Two things set the model apart at the API level. First, an agentic "thinking" pass: for complex prompts it plans composition and reasons through constraints before generating, which lifts success rates on infographics, multi-panel layouts, and text-heavy marketing assets. Second, an editing endpoint that accepts mask images for precise inpainting and outpainting, plus up to 16 reference images per call for identity and style consistency. Output is PNG at up to 2K native resolution, with neutral color that fixes the warm cast in gpt-image-1.5.
On GPTProto you call the same gpt-image-2 model ID through the standard OpenAI-compatible request shape — no separate OpenAI account, no organization verification step, and one balance that also covers Nano Banana Pro, Seedream, Flux, and 200+ other models.
| Spec | GPT Image 2 |
|---|---|
| Model ID | gpt-image-2 |
| Released | April 21, 2026 |
| Type | Text-to-image + image editing (inpaint / outpaint) |
| Output format | PNG (raster) |
| Max resolution | Native 2K (2048px); high-res variants up to ~4K |
| In-image text | Latin + CJK (Chinese / Japanese / Korean), dense layouts |
| Reasoning | Agentic "thinking mode" (plans before rendering) |
| Reference images | Up to 16 per call |
| Editing | Mask-based inpaint / outpaint; unedited pixels preserved |
| Input modality | Text (+ reference images) |
| OpenAI list price | $8 / $30 per 1M tokens (input / output) |
| GPTProto price | $6.4 / $24 per 1M tokens (20% under list) |
| Access | One GPTProto key, no OpenAI org verification |
| GPT Image 2 | Nano Banana Pro | GPT Image 1.5 | |
|---|---|---|---|
| Model ID | gpt-image-2 |
gemini-3-pro-image-preview |
gpt-image-1.5 |
| Vendor | OpenAI | OpenAI | |
| Released | Apr 2026 | Nov 2025 | Dec 2025 |
| Max resolution | 2K native (~4K variants ) | up to 4K (4096px) | 2K |
| In-image text | Latin + CJK | multilingual, long passages | improved vs gpt-image-1 |
| Reasoning | agentic thinking | Gemini 3 reasoning + Search grounding | none |
| Reference inputs | up to 16 images | multi-image, ~5-subject identity | fewer |
| OpenAI/Google list | $8 / $30 per 1M | $2 / $12 per 1M | $8 / $32 per 1M |
| GPTProto price | $6.4 / $24 per 1M | $ 0.0804/ per time |
$ 5.6 / $ 22.4 per 1M |
| On GPTProto | ✓ | ✓ | ✓ |
Which to pick (honest): Nano Banana Pro has the lower token rate, native 4K, and Search-grounded generation — reach for it when cost, 4K infographics, or real-world-grounded visuals matter most. GPT Image 2 leads on agentic layout planning, up to 16 reference images, and tight drop-in compatibility with the OpenAI SDK — reach for it for reference-heavy product/packaging work already wired to OpenAI. Because both run on one GPTProto balance, you can benchmark them against your own prompts without opening a second account.
If you already call gpt-image-2 on OpenAI, moving to GPTProto is a base-URL swap — the model ID and request shape stay the same:
model: "gpt-image-2" and your existing images.generate / edit request body. https://gptproto.com/v1 and use your GPTProto key.探索 gpt image 2 API 的技術細節和使用技巧,提升你的創作流程與影像品質。
與本模型相關的指南、對比與更新。
所有文章
比較 2026 年最便宜的 AI 影片生成器,了解每段影片的實際成本。Vidu Q3 Pro 每部影片僅需 $0.04,還有 Kling、Sora 2、Veo 及最佳免費工具。

Nano Banana 2 的費用是 Nano Banana Pro 的一半,且在 Image Arena 上的得分更高。了解每款 Gemini 影像模型各自勝出的情境,以及使用單一 API 執行兩者的程式碼。

Seedance 2.0 與 Kling 3.0 的人類動作表現比較:Kling 更擅長根據文字生成動作,Seedance 則更擅長複製參考動作。提供規格、盲測數據及可直接執行的 API 程式碼。

只需一張產品照片,就能以約 $25 製作 10 支原生 UGC 廣告。提供 Seedream 5.0 Pro + Seedance 2.0 API 的逐步工作流程,以及可直接執行的 Python 和 cURL。" slug: "ai-ugc-ads-seedream-5-pro-seedance-2
輸入
輸出
