🎁 新用戶 免費註冊,送 10 次呼叫,最高 $1,免綁卡。

gemini-3.1-flash-lite-image vs GPT Image 2

vs

何時該用哪一個 — 綜合評斷,而非基準測試數據表

兩者皆為圖片模型,每百萬輸出 token 皆收取相同的 $30,因此真正的差異在於輸入:gemini-3.1-flash-lite-image 接收文字與圖片為每百萬 $0.25,而 gpt-image-2 為 $5,使其在提供提示詞與參考圖片時的成本便宜了 20x。針對包含大量提示詞或圖片輸入並輸出圖片的管線,當輸入量主導帳單時請選擇 gemini-3.1-flash-lite-image,並請注意它也能在圖片之外回傳文字,具有 4096-token 的輸出上限,且無法關閉思考功能。當您需要來自 OpenAI 的純圖片輸出端點,且輸入成本只是次要的支出項目時,請選擇 gpt-image-2。

定價

gemini-3.1-flash-lite-image GPT Image 2 Δ
輸入 / 1M tokens $0.25 $5 0.05×
輸出 / 1M tokens $30 $30 =

費率取自建置時的即時目錄;各模型頁面皆附有目前的費率卡。

它們的相對位置 — 在此計費單位下,所有 8 個 圖像生成 模型的 每 1M tokens 的輸出價格(對數尺度)

規格

gemini-3.1-flash-lite-image GPT Image 2
輸入模態 文字 影像 文字 影像
輸出模態 文字 影像 影像
發布日期 2026-06-30 2026-04-21
知識截止日期 2025-01
輸出尺寸 1K only (1:1 = 1024x1024)
  • 1024x1024
  • 1536x1024
  • 1024x1536
  • 2048x2048
  • 2048x1152
  • 3840x2160
  • 2160x3840
  • auto
  • arbitrary WxH (both divisible by 16, aspect ratio 1:3–3:1)
輸入模式 text-to-image, interleaved generation + editing, up to 14 reference images text-to-image, image edit with mask inpainting
每次請求圖片數 10
格式 png, jpeg, webp
備註

1K (1024px) output only

10 aspect ratios (1:1, 3:2, 2:3, 3:4, 4:3, 4:5, 5:4, 9:16, 16:9, 21:9) and up to 14 reference images

interleaved generation and editing with fast multi-turn local edits

sub-2s end-to-end latency

SynthID + C2PA watermarking always on

Flexible resolutions: edges up to 3840px in multiples of 16, ratio <=3:1, ~0.65-8.3MP total (incl. 4K 3840x2160)

editing with mask inpainting

all image inputs processed at high fidelity

significantly improved text rendering (precise placement can still struggle)

規格摘錄自各供應商的文件;供應商未發布的資料列會直接省略,而非自行推測。 完整來源: gemini-3.1-flash-lite-image · GPT Image 2

單一提示詞,兩款模型 — 經由閘道測量

提示詞 A weathered enamel diner mug on a steel counter, the words "OPEN 24H" stencilled on the mug in worn paint, low winter sun raking in from the left, shallow depth of field.

gemini-3.1-flash-lite-image

gemini-3.1-flash-lite-image: A weathered enamel diner mug on a steel counter, the words "OPEN 24H" stencilled on the mug in worn paint, low winter sun raking in from the left, shallow depth of field.

模型回傳 1408×768 延遲 3 s

GPT Image 2

GPT Image 2: A weathered enamel diner mug on a steel counter, the words "OPEN 24H" stencilled on the mug in worn paint, low winter sun raking in from the left, shallow depth of field.

模型回傳 1402×1122 延遲 18 s

單一提示詞,每個模型一次請求,不重試且不擇優挑選——皆為各模型回傳的第一個結果。尺寸與長度皆未強制指定:每個模型均使用其預設值,因為若將請求調整為迎合所有模型,反而無法展現任何一方的優勢。此處的檔案皆已針對網頁重新編碼,因此請評斷其構圖與提示詞遵循度,而非壓縮品質。

只需一行程式碼即可在兩者間切換

以下每個頁籤中都有這兩個 ID — 醒目提示的這兩行是唯一的修改處。相同的端點,相同的金鑰,相同的請求結構。

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.images.generate(
    model="gemini-3.1-flash-lite-image",
    # model="gpt-image-2",  # 取消註解此行,並註解上一行
    prompt="a watercolor lighthouse at dawn",
    size="1024x1024",
)
print(resp.data[0].b64_json[:80])

取得 API 金鑰 →

常見問題

gemini-3.1-flash-lite-image 和 GPT Image 2 哪個比較便宜?

gemini-3.1-flash-lite-image 在 輸入 / 1m tokens 上較便宜($0.25 對比 $5,相差 20×)。其他項目可能呈現相反結果 — 上表提供完整資訊,實際成本取決於您的使用組合。

我可以在不進行兩次整合的情況下,對 gemini-3.1-flash-lite-image 和 GPT Image 2 進行 A/B 測試嗎?

可以。兩者皆透過同一個相容 OpenAI 的端點提供服務,並使用同一把 API 金鑰 — 切換只需更改一行的模型字串,因此您可以將部分流量分別導向兩者並直接比較帳單。

價格會隨圖片大小改變嗎?

這取決於模型的計費方式。按圖片計費的模型無論提示或輸出大小為何,收費皆相同;按 token 計費的模型則會根據您生成的解析度等比例縮放,因此 4K 圖片的成本會是小圖片的數倍。上表顯示了各模型適用的計費方式。

相關比較