新人 免费注册,送 10 次调用,最高 $1,免绑卡。

GPT-6.1 Sol

发布于 2026-09-29

chat图像理解代码工具调用推理提示词缓存

GPT-6.1 Sol 是 GPT-6 Sol 的升级版,2026 年 9 月 29 日发布,OpenAI 将它定位为以更低成本提供接近 Astra 的表现,面向复杂编码、计算机操作与专业工作。

输入
文本 图像 $2/M
输出
文本 $10/M
缓存读取
$0.1/M
上下文
1.1M
对比 GPT-4o
便宜约 60%
知识截止
2026-04

输入超过 272K tokens 时,整单按输入 $4/M、输出 $15/M 计价

价格在同类中的位置

价格在 68 个同类模型中的位置

输入$2/M
$0.05 · Qwen3 VL Flash GPT-5.4 Pro · $30
输出$10/M
$0.275 · DeepSeek V4 Flash GPT-5.4 Pro · $180
缓存读取$0.1/M
$0.0028 · DeepSeek V4 Flash GPT-5.4 Pro · $15

这条线显示该模型的价格,在 Synthorai 上同类模型里处于什么位置。两端标出了最便宜和最贵的那个。这里是基础价,批量、区域和缓存写入的折扣见价格页。

规格与限制

Token

上下文窗口(厂商规格) 1,050,000
最大输出(厂商规格) 128,000
知识截止 2026-04

提示词缓存

缓存方式 自动
最低前缀 1,024
存活时间 5–10 分钟,最长 1 小时

思考

厂商参数 reasoning.effort
可选值 low · medium · high · xhigh · max
默认值 medium 请求未指定时生效
可关闭 不支持
参数 reasoning_effort
取值 minimal · low · medium · high 网关侧参数面——以上方厂商映射为准

模型

模态 文本 + 图像 → 文本
  • Upgrade to GPT-6 Sol with near-Astra performance for complex coding, computer use and professional work. 1.05M context of which up to 922K can be input, 128K max output. Cached input is 0.05x the input rate, half of GPT-6 Sol's. Prompts over 272K input tokens move the whole request onto the long-context rate (2x input and cache, 1.5x output). Reasoning cannot be turned off: none and minimal are not supported. Tool calling requires the Responses API
  • Chat Completions is supported without tools

依据 OpenAI 官方文档 ↗

单个 Prompt —— 通过网关实测

提示词 Explain to a CFO, in exactly three sentences, why our LLM bill doubled after we enabled reasoning. 检查 恰好 3 句

GPT-6.1 Sol 通过 · 3 sentences

Enabling reasoning makes the model do more computational work before answering, and many providers bill the resulting reasoning tokens even though users never see them. That means the same number of requests can cost twice as much because each request consumes more billable tokens, even if the visible answers stay the same length. We should confirm the increase in our usage breakdown and reserve higher reasoning levels for tasks where the improvement justifies the added cost.

输出 144 tok (+54 思考) 延迟 6.6 s

指令遵循(恰好三句,可数)、受众适配(面向 CFO 的语气),以及下方 token 计量所暴露的隐藏思考计费缺口。

提示词 This function is supposed to return the longest run of consecutive equal items, but callers report it is off by one on single-element inputs and misses runs that end at the last index. Fix it and explain the bug in one paragraph. 检查 修复通过测试

GPT-6.1 Sol 通过 · 8/8 cases

```python def longest_run(xs): if not xs: return 0 best = cur = 1 for i in range(1, len(xs)): if xs[i] == xs[i - 1]: cur += 1 else: cur = 1 best = max(best, cur) return best ``` The original function counts equal adjacent pairs rather than items in a run: a run of length `n` has only `n - 1` equal pairs, so a single-element input returns `0`, and a run ending at the last index is also undercounted. It does process the last index, so the issue is not a missing final check. Initializing each nonempty run to `1` and updating `best` after every iteration correctly counts items, including singleton and trailing runs; empty inputs still return `0`.

输出 189 tok 延迟 6.3 s

修复是否真的正确(可运行)、解释的信息密度,以及在一个边界明确的任务上的 token 效率。

提示词 Extract a JSON object with fields {company, quarter, revenue_usd_m, growth_pct, guidance} from this text. Use null for anything not stated; add no other fields. 检查 合法 JSON,schema 精确

GPT-6.1 Sol 通过 · 5/5 fields, guidance "Withheld pending the Q4 close"

{ "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": null, "growth_pct": "eleven and change", "guidance": "Withheld pending the Q4 close" }

输出 345 tok (+285 思考) 延迟 10.2 s

schema 服从度(不臆造字段)、幻觉压力(guidance 明确被暂缓给出),以及结构化输出路径的差异。

提示词 Write a 120-word product announcement for an API gateway feature that caches prompts across providers. Forbidden words: "seamless", "unlock", "game-changer", "revolutionize", "empower". Exactly one sentence must be a question. 检查 120 词,0 个禁用词

GPT-6.1 Sol 通过 · 120 words, 0 banned, 1 question

Introducing Cross-Provider Prompt Cache, a new API gateway feature that stores reusable prompts and manages caching across your supported AI providers. Why rebuild the same context every time your application switches models? With one configuration, teams can reuse shared instructions, standardize cache policies, and reduce repeated prompt processing wherever provider caching is available. The gateway handles provider-specific requirements while giving you clear visibility into cache hits, usage, and estimated savings. Set expiration windows, isolate cached content by project, and invalidate entries when prompts change. Your existing routing logic stays intact, so you can compare models without rebuilding your caching workflow. […]

输出 588 tok (+435 思考) 延迟 13.9 s

约束服从度(字数预算、禁用词表、唯一的那句问句)、文风指纹,以及长度控制。

30 秒用上 GPT-6.1 Sol

OpenAI 兼容:换掉 base_url,SDK 不用改。POST /v1/chat/completions

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.chat.completions.create(
    model="gpt-6.1-sol",
    messages=[{"role": "user", "content": "Summarize this diff"}],
    reasoning_effort="medium",
)
print(resp.choices[0].message.content)

关于 GPT-6.1 Sol

  • 它沿用 GPT-6 家族的 1,050,000 token 上下文窗口(其中输入最多 922,000 token)与 128K 最大输出,支持文本与图像输入、结构化输出、流式、函数调用和提示词缓存,知识截止为 2026 年 4 月。
  • 输入与输出的刊例价与 GPT-6 Sol 相同,但缓存输入价减半,对反复重发长前缀的智能体循环来说积少成多。
  • 输入超过 272K token 的提示词,整个请求转入长上下文价,输入与缓存为两倍、输出为 1.5 倍。
  • 推理始终开启:力度从 low 到 max,默认 medium,OpenAI 在这个模型上不支持 none 与 minimal,因此 Synthorai 会把要求这两档的请求按 low 执行,而不是拒绝。
  • OpenAI 规定函数调用须走 Responses API;通过 Synthorai,带工具的 Chat Completions 请求依然可用,因为网关会以支持函数调用的形式发往上游。
  • Synthorai 通过与其余模型相同的 OpenAI 兼容 API 提供 GPT-6.1 Sol。

常见问题

GPT-6.1 Sol API 可以免费试用吗?

可以。新账号可获得 10 次试用调用和最高 $1 的免费额度,无需绑卡。按输入 $2/M 计算,仅这笔额度就足够对 GPT-6.1 Sol 发起约 62 次 ~8K token 的请求。

GPT-6.1 Sol 最擅长什么?

GPT-6 Sol 的升级版,编码、计算机操作与专业工作接近 Astra、1.05M 上下文、128K 输出,推理力度 low 到 max、缓存输入价为 GPT-6 Sol 的一半,272K 输入以上按长上下文计价。完整能力请见「关于」部分,内容取自厂商官方发布说明。

GPT-6.1 Sol 的价格是多少?

在 Synthorai 上,GPT-6.1 Sol 输入 $2/百万 token、输出 $10/百万 token,即厂商牌价,无平台加价。缓存命中的输入 token 按 $0.1/M 计费。

GPT-6.1 Sol 支持提示词缓存(prompt caching)吗?

支持,且全自动:经 OpenAI 提供的提示词自动缓存,无需改代码。缓存命中的输入 token 按 $0.1/M 计费(未命中 $2/M);提示词需有 1,024 token 以上的稳定前缀才能命中缓存(TTL 5–10 分钟,最长 1 小时)。 提示词缓存指南 →

如何开通 GPT-6.1 Sol?

把现有 OpenAI SDK 的 base_url 指向 "https://synthorai.io/v1",model 设为 "gpt-6.1-sol" 即可。一把 API key 通用网关上的全部模型。

GPT-6.1 Sol 的知识截止日期是什么时候?

GPT-6.1 Sol 的知识截止日期为 2026-04,依据厂商官方文档(数据核验于 2026-09-30)。

相关模型

对比

本页每个值都转录自厂商自己的文档(链接见上),并带有核对日期。价格在全目录范围内比较;各厂商定义不同的规格值,只说明差异而不作图表对比。此处没有任何由我们测量的数据,也不做评分。

获取 API 密钥 算算你的成本 →