Claude Sonnet 5.5 是 Anthropic 在速度与智能之间取得最佳平衡的模型,于 2026 年 9 月 28 日发布,价格与 Claude Sonnet 5 相同:每百万输入 token $2、每百万输出 token $10,提示词缓存读取为输入价的 10%(每百万 $0.20),5 分钟缓存写入 $2.50,1 小时缓存写入 $4。
- 输入
- 文本 图像 $2/M
- 输出
- 文本 $10/M
- 缓存读取
- $0.2/M
- 上下文
- 1M
- 对比 GPT-4o
- 便宜约 60%
- 知识截止
- 2026-06
价格在同类中的位置
价格在 68 个同类模型中的位置
这条线显示该模型的价格,在 Synthorai 上同类模型里处于什么位置。两端标出了最便宜和最贵的那个。这里是基础价,批量、区域和缓存写入的折扣见价格页。
规格与限制
Token
| 上下文窗口(厂商规格) | 1,000,000 |
|---|---|
| 最大输出(厂商规格) | 128,000 |
| 知识截止 | 2026-06 |
提示词缓存
| 缓存方式 | 显式(需开启) |
|---|---|
| 最低前缀 | 512 厂商默认 1,024 |
| 存活时间 | 默认 5 分钟,可选 1 小时 |
| 写入成本 | 1.25x (5m) / 2x (1h) |
思考
| 厂商参数 | thinking.type |
|---|---|
| 可选值 | adaptive (default) · between_tools |
| 默认值 | adaptive, effort high 请求未指定时生效 |
| 可关闭 | 不支持 |
| 思考行为 | Adaptive thinking is on by default. The lowest setting, between_tools, turns off up-front thinking and works at high effort or below; thinking {"type": "disabled"} and a manual {"type": "enabled", "budget_tokens": N} both return a 400 error. |
| 参数 | reasoning_effort |
| 取值 | minimal · low · medium · high 网关侧参数面——以上方厂商映射为准 |
模型
| 模态 | 文本 + 图像 → 文本 |
|---|
- Same price as Claude Sonnet 5
- 1M context at standard pricing with no long-context tier
- prompt-cache reads cost 0.1x input ($0.20/M)
- minimum cacheable prompt 512 tokens
- setting temperature, top_p or top_k to a non-default value returns a 400
- forced tool use returns an error
- text between tool calls comes back in thinking blocks
- up to 300k output tokens on the Message Batches API with the output-300k-2026-03-24 beta header
单个 Prompt —— 通过网关实测
Claude Sonnet 5.5 通过 · 3 sentences
Enabling reasoning makes the model "think" before it answers, and that internal thinking is generated as extra text (reasoning tokens) that we are billed for at the same rate as output tokens, even though users never see it. Output tokens also cost several times more than input tokens, so a single request that used to produce a short answer can now generate many times more billable text behind the scenes. Because our request volume and prompts didn't change, the doubling comes almost entirely from this hidden thinking, and we can bring it down by limiting reasoning effort, turning it off for simple tasks, and reserving it for the workloads where the accuracy gain justifies the cost.
输出 205 tok 延迟 3.6 s
指令遵循(恰好三句,可数)、受众适配(面向 CFO 的语气),以及下方 token 计量所暴露的隐藏思考计费缺口。
Claude Sonnet 5.5 通过 · 8/8 cases
**The bug:** `cur` counts the number of *adjacent equal pairs* rather than the number of items in the run. A run of length k contains only k−1 such pairs, so every result is one too small. The counter also starts at 0 and the loop starts at index 1, so a single-element list never updates `best` and returns 0 instead of 1. (An empty list should return 0.) In addition, `cur` resets to 0 on a mismatch, when it should reset to 1, because the item at `xs[i]` begins a new run of length 1. The claim that it "misses runs ending at the last index" is really the same off-by-one: the final run is counted, but as one less than its true length. Because `best` is updated inside the loop on every extension, the last run isn't skipped. […]
输出 444 tok 延迟 6.7 s
修复是否真的正确(可运行)、解释的信息密度,以及在一个边界明确的任务上的 token 效率。
Claude Sonnet 5.5 通过 · 5/5 fields, guidance null
Here is the extracted JSON: ```json { "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": 412, "growth_pct": 11, "guidance": null } ``` Notes on the choices: - **quarter**: The text says "September quarter" without specifying a fiscal year or Q-number, so I kept the wording as stated. - **revenue_usd_m**: The text gives "412" for the topline, with no currency or unit stated explicitly. I've taken it as 412 million USD per the field definition, but the text does not confirm that, and the figure is constant-currency. - **growth_pct**: "Eleven and change" means slightly above 11, so 11 is an approximation. The exact figure isn't given. - **guidance**: Guidance is withheld pending the Q4 close, so no guidance is stated, hence null.
输出 300 tok 延迟 3.5 s
schema 服从度(不臆造字段)、幻觉压力(guidance 明确被暂缓给出),以及结构化输出路径的差异。
Claude Sonnet 5.5 通过 · 122 words, 0 banned, 1 question
**Introducing Cross-Provider Prompt Caching** Your API gateway now caches prompts across every connected LLM provider, whether you route traffic to OpenAI, Anthropic, Google, or a self-hosted model. Here's how it works: the gateway fingerprints each prompt prefix, stores it centrally, and applies the right provider-specific caching mechanism automatically. When you switch providers or fail over during an outage, your cached context follows you. Early testing shows input token costs dropping by up to 70% on repeated system prompts and long documents, with latency improvements of 40% or more on cache hits. Tired of paying to send the same instructions thousands of times a day? Enable it with one configuration flag. Cache rules, TTLs, and per-team isolation are available in the dashboard today.
输出 266 tok 延迟 3.8 s
约束服从度(字数预算、禁用词表、唯一的那句问句)、文风指纹,以及长度控制。
30 秒用上 Claude Sonnet 5.5
OpenAI 兼容:换掉 base_url,SDK 不用改。POST /v1/chat/completions
from openai import OpenAI
client = OpenAI(
base_url="https://synthorai.io/v1",
api_key="sk-syn-...",
)
resp = client.chat.completions.create(
model="claude-sonnet-5-5",
messages=[{"role": "user", "content": "Summarize this diff"}],
reasoning_effort="medium",
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://synthorai.io/v1",
apiKey: "sk-syn-...",
});
const resp = await client.chat.completions.create({
model: "claude-sonnet-5-5",
messages: [{ role: "user", content: "Summarize this diff" }],
reasoning_effort: "medium",
});
console.log(resp.choices[0].message.content);curl https://synthorai.io/v1/chat/completions \
-H "Authorization: Bearer sk-syn-..." \
-H "Content-Type: application/json" \
-d '{
"model": "claude-sonnet-5-5",
"messages": [{"role": "user", "content": "Hello"}],
"reasoning_effort": "medium"
}'package main
import (
"context"
"fmt"
"github.com/openai/openai-go/v3"
"github.com/openai/openai-go/v3/option"
)
func main() {
client := openai.NewClient(
option.WithBaseURL("https://synthorai.io/v1"),
option.WithAPIKey("sk-syn-..."),
)
resp, _ := client.Chat.Completions.New(context.TODO(), openai.ChatCompletionNewParams{
Model: "claude-sonnet-5-5",
Messages: []openai.ChatCompletionMessageParamUnion{
openai.UserMessage("Summarize this diff"),
},
ReasoningEffort: openai.ReasoningEffortMedium,
})
fmt.Println(resp.Choices[0].Message.Content)
}import com.openai.client.OpenAIClient;
import com.openai.client.okhttp.OpenAIOkHttpClient;
import com.openai.models.chat.completions.*;
import com.openai.models.ReasoningEffort;
OpenAIClient client = OpenAIOkHttpClient.builder()
.baseUrl("https://synthorai.io/v1")
.apiKey("sk-syn-...")
.build();
ChatCompletion resp = client.chat().completions().create(
ChatCompletionCreateParams.builder()
.model("claude-sonnet-5-5")
.addUserMessage("Summarize this diff")
.reasoningEffort(ReasoningEffort.MEDIUM)
.build());
System.out.println(resp.choices().get(0).message().content().orElse(""));关于 Claude Sonnet 5.5
- 它提供 1M token 上下文窗口且无长上下文加价,最大输出 128K,支持文本与图像输入。
- 自适应思考默认开启,默认强度为 high;最低档 between_tools 会关闭「先想后答」,在 high 及以下强度可用,而把 thinking 设为 disabled 或手动指定 budget_tokens 都会返回 400。
- temperature、top_p、top_k 设为非默认值同样返回 400,强制工具调用也不受支持。
- Anthropic 列出了五项会影响已在 Claude Sonnet 5 上运行代码的破坏性变更,工具调用之间的文字现在会以 thinking 块返回。
- 可缓存提示词的最小长度为 512 token。
- Synthorai 通过与其余模型相同的 OpenAI 兼容 API 提供 Claude Sonnet 5.5。
常见问题
Claude Sonnet 5.5 API 可以免费试用吗?
可以。新账号可获得 10 次试用调用和最高 $1 的免费额度,无需绑卡。按输入 $2/M 计算,仅这笔额度就足够对 Claude Sonnet 5.5 发起约 62 次 ~8K token 的请求。
Claude Sonnet 5.5 最擅长什么?
当前阵容中速度与智能平衡最佳、与 Sonnet 5 同价:每百万输入 $2、输出 $10、between_tools 可关闭「先想后答」。完整能力请见「关于」部分,内容取自厂商官方发布说明。
Claude Sonnet 5.5 的价格是多少?
在 Synthorai 上,Claude Sonnet 5.5 输入 $2/百万 token、输出 $10/百万 token,即厂商牌价,无平台加价。缓存命中的输入 token 按 $0.2/M 计费。
Claude Sonnet 5.5 支持提示词缓存(prompt caching)吗?
支持,需主动开启:用 cache_control 断点标记稳定前缀。缓存命中的输入 token 按 $0.2/M 计费(未命中 $2/M);提示词需有 512 token 以上的稳定前缀才能命中缓存(TTL 默认 5 分钟,可选 1 小时)。 提示词缓存指南 →
如何开通 Claude Sonnet 5.5?
把现有 OpenAI SDK 的 base_url 指向 "https://synthorai.io/v1",model 设为 "claude-sonnet-5-5" 即可。一把 API key 通用网关上的全部模型。
Claude Sonnet 5.5 的知识截止日期是什么时候?
Claude Sonnet 5.5 的知识截止日期为 2026-06,依据厂商官方文档(数据核验于 2026-09-29)。
相关模型
对比
本页每个值都转录自厂商自己的文档(链接见上),并带有核对日期。价格在全目录范围内比较;各厂商定义不同的规格值,只说明差异而不作图表对比。此处没有任何由我们测量的数据,也不做评分。