🎁 新人 免费注册,送 10 次调用,最高 $1,免绑卡。

gemini-3.1-flash-lite-image vs wan2.7-image-pro

vs

何时使用哪一个 — 经过整理的结论,而非基准测试表格

这两者的计费单位不同——gemini-3.1-flash-lite-image 按 token 计费(输入 $0.25 每百万,输出 $30 每百万),而 wan2.7-image-pro 按次调用计费(每次调用 $0.075)——因此任何单一的换算方式都是不准确的,你应该根据自己平均的请求形态进行比较。当你希望从文本或图像输入中同时获得返回的文本和图像,并且输出上限为 4096 token、知识截止日期为 2025-01 时,请选择 gemini-3.1-flash-lite-image。对于纯图像输出生成,请选择 wan2.7-image-pro,其固定的每次调用计费使得无论提示词长度如何,成本都可预测。

定价

gemini-3.1-flash-lite-image wan2.7-image-pro Δ
每张生成图像 $0.075
输入 / 1M tokens $0.25
输出 / 1M tokens $30

这两个模型的计费单位不同,因此不显示 Δ — 它们之间的转换需要基于我们尚未实测的假设。上面列出的每张费率卡均采用其自身对应的单位。

规格

gemini-3.1-flash-lite-image wan2.7-image-pro
输入模态 文本 图像 文本 图像
输出模态 文本 图像 图像
发布日期 2026-06-30 2026-04-01
知识截止日期 2025-01
输出尺寸 1K only (1:1 = 1024x1024)
  • 1K (1024x1024)
  • 2K (2048x2048, default)
  • 4K (4096x4096, text-to-image only)
  • custom WxH (t2i 768x768 – 4096x4096; editing/sets 768x768 – 2048x2048, ratio 1:8–8:1)
输入模式 text-to-image, interleaved generation + editing, up to 14 reference images text-to-image, image editing (incl. bounding-box interactive edit), 0–9 reference images, text/image-to-image-set
每次请求图像数 12
格式 png
备注

1K (1024px) output only

10 aspect ratios (1:1, 3:2, 2:3, 3:4, 4:3, 4:5, 5:4, 9:16, 16:9, 21:9) and up to 14 reference images

interleaved generation and editing with fast multi-turn local edits

sub-2s end-to-end latency

SynthID + C2PA watermarking always on

4K text-to-image (max 4096x4096; editing max 2K), aspect ratios 1:8-8:1

instruction + click-to-edit editing

character-consistent sets up to 12 images

print-quality text rendering incl. formulas/tables

规格转录自各供应商的文档;若供应商未发布某项数据,则直接省略该行,而非进行推断。 完整来源: gemini-3.1-flash-lite-image · wan2.7-image-pro

单个 Prompt,两个模型 —— 通过网关实测

提示词 A weathered enamel diner mug on a steel counter, the words "OPEN 24H" stencilled on the mug in worn paint, low winter sun raking in from the left, shallow depth of field.

gemini-3.1-flash-lite-image

gemini-3.1-flash-lite-image: A weathered enamel diner mug on a steel counter, the words "OPEN 24H" stencilled on the mug in worn paint, low winter sun raking in from the left, shallow depth of field.

模型返回 1408×768 延迟 3 s

wan2.7-image-pro

wan2.7-image-pro: A weathered enamel diner mug on a steel counter, the words "OPEN 24H" stencilled on the mug in worn paint, low winter sun raking in from the left, shallow depth of field.

模型返回 1024×1024 延迟 25 s

统一提示词,每个模型单次请求,无重试且无择优挑选 —— 均为各模型返回的首个结果。尺寸与时长均未固定:各模型使用自身的默认值,因为试图适配所有模型的请求参数无法展现任何模型的优势。此处文件已为 Web 重新编码,因此请评判构图与提示词遵循度,而非压缩质量。

只需一行代码即可在它们之间切换

两个 ID 都包含在下方的每个选项卡中 — 高亮显示的两行是唯一的修改。相同的端点,相同的密钥,相同的请求结构。

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.images.generate(
    model="gemini-3.1-flash-lite-image",
    # model="wan2.7-image-pro",  # 取消注释此行,注释上一行
    prompt="a watercolor lighthouse at dawn",
    size="1024x1024",
)
print(resp.data[0].b64_json[:80])

获取 API 密钥 →

常见问题

gemini-3.1-flash-lite-image 和 wan2.7-image-pro 哪个更便宜?

它们的计费单位不同,因此不存在单一的绝对数字:在上方表格中,gemini-3.1-flash-lite-image 和 wan2.7-image-pro 分别以各自的单位显示。请根据你自己的工作负载进行比较——本页顶部的结论中描述了实际的权衡。

我可以在不进行两次集成的情况下,对 gemini-3.1-flash-lite-image 和 wan2.7-image-pro 进行 A/B 测试吗?

可以。两者均通过同一个兼容 OpenAI 的端点提供服务,并使用同一把 API 密钥——切换只需更改一行模型字符串,因此你可以将一部分流量路由到各个模型并直接比较账单。

价格会随图像尺寸变化吗?

这取决于模型的计费方式。按图像计费的模型,无论提示词或输出尺寸如何,收费均相同;按 token 计费的模型则会随着你渲染的分辨率而增加费用,因此 4K 图像的成本是小尺寸图像的数倍。上方表格标明了各个模型适用的计费方式。

相关对比