Claude Sonnet 5.5 是 Anthropic 在速度與智慧之間取得最佳平衡的模型,於 2026 年 9 月 28 日發布,價格與 Claude Sonnet 5 相同:每百萬輸入 token $2、每百萬輸出 token $10,提示詞快取讀取為輸入價的 10%(每百萬 $0.20),5 分鐘快取寫入 $2.50,1 小時快取寫入 $4。
- 輸入
- 文字 影像 $2/M
- 輸出
- 文字 $10/M
- 快取讀取
- $0.2/M
- 上下文
- 1M
- 相較 GPT-4o
- 便宜約 60%
- 知識截止
- 2026-06
價格在同類中的位置
價格在 68 個同類模型中的位置
這條線顯示該模型的價格,在 Synthorai 上同類模型裡處於什麼位置。兩端標出了最便宜和最貴的那個。這裡是基礎價,批次、區域與快取寫入的折扣見價格頁。
規格與限制
Token
| 上下文視窗(廠商規格) | 1,000,000 |
|---|---|
| 最大輸出(廠商規格) | 128,000 |
| 知識截止 | 2026-06 |
提示詞快取
| 快取方式 | 顯式(需開啟) |
|---|---|
| 最低前綴 | 512 廠商預設 1,024 |
| 存活時間 | 預設 5 分鐘,可選 1 小時 |
| 寫入成本 | 1.25x (5m) / 2x (1h) |
思考
| 廠商參數 | thinking.type |
|---|---|
| 可選值 | adaptive (default) · between_tools |
| 預設值 | adaptive, effort high 請求未指定時生效 |
| 可關閉 | 不支援 |
| 思考行為 | Adaptive thinking is on by default. The lowest setting, between_tools, turns off up-front thinking and works at high effort or below; thinking {"type": "disabled"} and a manual {"type": "enabled", "budget_tokens": N} both return a 400 error. |
| 參數 | reasoning_effort |
| 取值 | minimal · low · medium · high 閘道側參數面——以上方廠商映射為準 |
模型
| 模態 | 文字 + 影像 → 文字 |
|---|
- Same price as Claude Sonnet 5
- 1M context at standard pricing with no long-context tier
- prompt-cache reads cost 0.1x input ($0.20/M)
- minimum cacheable prompt 512 tokens
- setting temperature, top_p or top_k to a non-default value returns a 400
- forced tool use returns an error
- text between tool calls comes back in thinking blocks
- up to 300k output tokens on the Message Batches API with the output-300k-2026-03-24 beta header
單一提示詞 — 經由閘道測量
Claude Sonnet 5.5 通過 · 3 sentences
Enabling reasoning makes the model "think" before it answers, and that internal thinking is generated as extra text (reasoning tokens) that we are billed for at the same rate as output tokens, even though users never see it. Output tokens also cost several times more than input tokens, so a single request that used to produce a short answer can now generate many times more billable text behind the scenes. Because our request volume and prompts didn't change, the doubling comes almost entirely from this hidden thinking, and we can bring it down by limiting reasoning effort, turning it off for simple tasks, and reserving it for the workloads where the accuracy gain justifies the cost.
輸出 205 tok 延遲 3.6 s
指令遵循(恰好三句,可數)、受眾適配(面向 CFO 的語氣),以及下方 token 計量所暴露的隱藏思考計費缺口。
Claude Sonnet 5.5 通過 · 8/8 cases
**The bug:** `cur` counts the number of *adjacent equal pairs* rather than the number of items in the run. A run of length k contains only k−1 such pairs, so every result is one too small. The counter also starts at 0 and the loop starts at index 1, so a single-element list never updates `best` and returns 0 instead of 1. (An empty list should return 0.) In addition, `cur` resets to 0 on a mismatch, when it should reset to 1, because the item at `xs[i]` begins a new run of length 1. The claim that it "misses runs ending at the last index" is really the same off-by-one: the final run is counted, but as one less than its true length. Because `best` is updated inside the loop on every extension, the last run isn't skipped. […]
輸出 444 tok 延遲 6.7 s
修復是否真的正確(可執行)、解釋的資訊密度,以及在一個邊界明確的任務上的 token 效率。
Claude Sonnet 5.5 通過 · 5/5 fields, guidance null
Here is the extracted JSON: ```json { "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": 412, "growth_pct": 11, "guidance": null } ``` Notes on the choices: - **quarter**: The text says "September quarter" without specifying a fiscal year or Q-number, so I kept the wording as stated. - **revenue_usd_m**: The text gives "412" for the topline, with no currency or unit stated explicitly. I've taken it as 412 million USD per the field definition, but the text does not confirm that, and the figure is constant-currency. - **growth_pct**: "Eleven and change" means slightly above 11, so 11 is an approximation. The exact figure isn't given. - **guidance**: Guidance is withheld pending the Q4 close, so no guidance is stated, hence null.
輸出 300 tok 延遲 3.5 s
schema 服從度(不臆造欄位)、幻覺壓力(guidance 明確被暫緩給出),以及結構化輸出路徑的差異。
Claude Sonnet 5.5 通過 · 122 words, 0 banned, 1 question
**Introducing Cross-Provider Prompt Caching** Your API gateway now caches prompts across every connected LLM provider, whether you route traffic to OpenAI, Anthropic, Google, or a self-hosted model. Here's how it works: the gateway fingerprints each prompt prefix, stores it centrally, and applies the right provider-specific caching mechanism automatically. When you switch providers or fail over during an outage, your cached context follows you. Early testing shows input token costs dropping by up to 70% on repeated system prompts and long documents, with latency improvements of 40% or more on cache hits. Tired of paying to send the same instructions thousands of times a day? Enable it with one configuration flag. Cache rules, TTLs, and per-team isolation are available in the dashboard today.
輸出 266 tok 延遲 3.8 s
約束服從度(字數預算、禁用詞表、唯一的那句問句)、文風指紋,以及長度控制。
30 秒用上 Claude Sonnet 5.5
OpenAI 相容:換掉 base_url,SDK 不用改。POST /v1/chat/completions
from openai import OpenAI
client = OpenAI(
base_url="https://synthorai.io/v1",
api_key="sk-syn-...",
)
resp = client.chat.completions.create(
model="claude-sonnet-5-5",
messages=[{"role": "user", "content": "Summarize this diff"}],
reasoning_effort="medium",
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://synthorai.io/v1",
apiKey: "sk-syn-...",
});
const resp = await client.chat.completions.create({
model: "claude-sonnet-5-5",
messages: [{ role: "user", content: "Summarize this diff" }],
reasoning_effort: "medium",
});
console.log(resp.choices[0].message.content);curl https://synthorai.io/v1/chat/completions \
-H "Authorization: Bearer sk-syn-..." \
-H "Content-Type: application/json" \
-d '{
"model": "claude-sonnet-5-5",
"messages": [{"role": "user", "content": "Hello"}],
"reasoning_effort": "medium"
}'package main
import (
"context"
"fmt"
"github.com/openai/openai-go/v3"
"github.com/openai/openai-go/v3/option"
)
func main() {
client := openai.NewClient(
option.WithBaseURL("https://synthorai.io/v1"),
option.WithAPIKey("sk-syn-..."),
)
resp, _ := client.Chat.Completions.New(context.TODO(), openai.ChatCompletionNewParams{
Model: "claude-sonnet-5-5",
Messages: []openai.ChatCompletionMessageParamUnion{
openai.UserMessage("Summarize this diff"),
},
ReasoningEffort: openai.ReasoningEffortMedium,
})
fmt.Println(resp.Choices[0].Message.Content)
}import com.openai.client.OpenAIClient;
import com.openai.client.okhttp.OpenAIOkHttpClient;
import com.openai.models.chat.completions.*;
import com.openai.models.ReasoningEffort;
OpenAIClient client = OpenAIOkHttpClient.builder()
.baseUrl("https://synthorai.io/v1")
.apiKey("sk-syn-...")
.build();
ChatCompletion resp = client.chat().completions().create(
ChatCompletionCreateParams.builder()
.model("claude-sonnet-5-5")
.addUserMessage("Summarize this diff")
.reasoningEffort(ReasoningEffort.MEDIUM)
.build());
System.out.println(resp.choices().get(0).message().content().orElse(""));關於 Claude Sonnet 5.5
- 它提供 1M token 上下文視窗且無長上下文加價,最大輸出 128K,支援文字與圖像輸入。
- 自適應思考預設開啟,預設強度為 high;最低檔 between_tools 會關閉「先想後答」,在 high 及以下強度可用,而把 thinking 設為 disabled 或手動指定 budget_tokens 都會回傳 400。
- temperature、top_p、top_k 設為非預設值同樣回傳 400,強制工具呼叫也不受支援。
- Anthropic 列出了五項會影響已在 Claude Sonnet 5 上執行程式碼的破壞性變更,工具呼叫之間的文字現在會以 thinking 區塊回傳。
- 可快取提示詞的最小長度為 512 token。
- Synthorai 透過與其餘模型相同的 OpenAI 相容 API 提供 Claude Sonnet 5.5。
常見問題
Claude Sonnet 5.5 API 可以免費試用嗎?
可以,新帳號可獲得 10 次試用呼叫和最高 $1 的免費額度,無需信用卡。以輸入 $2/M 計算,光是這筆額度就足以對 Claude Sonnet 5.5 發出約 62 次 ~8K token 的請求。
Claude Sonnet 5.5 最擅長什麼?
目前陣容中速度與智慧平衡最佳、與 Sonnet 5 同價:每百萬輸入 $2、輸出 $10、between_tools 可關閉「先想後答」。完整能力請見「關於」一節,內容取自廠商官方發布說明。
Claude Sonnet 5.5 的價格是多少?
在 Synthorai 上,Claude Sonnet 5.5 輸入 $2/百萬 token、輸出 $10/百萬 token,即廠商牌價,無平台加價。快取命中的輸入 token 以 $0.2/M 計費。
Claude Sonnet 5.5 支援提示詞快取(prompt caching)嗎?
支援,需主動開啟:以 cache_control 中斷點標記穩定前綴。快取命中的輸入 token 以 $0.2/M 計費(未命中 $2/M);提示詞需有 512 個 token 以上的穩定前綴才能命中快取(TTL 預設 5 分鐘,可選 1 小時)。 提示詞快取指南 →
如何開通 Claude Sonnet 5.5?
把現有 OpenAI SDK 的 base_url 指向 "https://synthorai.io/v1",model 設為 "claude-sonnet-5-5" 即可。一組 API key 通用閘道上的所有模型。
Claude Sonnet 5.5 的知識截止日期是什麼時候?
Claude Sonnet 5.5 的知識截止日期為 2026-06,依據廠商官方文件(資料核驗於 2026-09-29)。
相關模型
對比
本頁每個值都轉錄自廠商自己的文件(連結見上),並帶有核對日期。價格在全目錄範圍內比較;各廠商定義不同的規格值,只說明差異而不作圖表對比。此處沒有任何由我們測量的資料,也不做評分。