Claude Sonnet 5.5 は速度と知性のバランスが最も優れた Anthropic のモデルで、2026 年 9 月 28 日にリリースされ、価格は Claude Sonnet 5 と同じです。
- 入力
- テキスト 画像 $2/M
- 出力
- テキスト $10/M
- キャッシュ読み取り
- $0.2/M
- コンテキスト
- 1M
- GPT-4o 比
- 約 60% 割安
- 知識カットオフ
- 2026-06
価格の位置づけ
同種 68 モデル中の料金の位置
このバーは、Synthorai 上の同種モデルの中でこのモデルの価格がどこに位置するかを示します。両端には最も安いモデルと最も高いモデルの名前が入ります。表示は基本料金で、バッチ・リージョン・キャッシュ書き込みの割引は料金ページにあります。
スペックと制限
トークン
| コンテキストウィンドウ(ベンダー仕様) | 1,000,000 |
|---|---|
| 最大出力(ベンダー仕様) | 128,000 |
| 知識カットオフ | 2026-06 |
プロンプトキャッシュ
| キャッシュ方式 | 明示(オプトイン) |
|---|---|
| 最小プレフィックス | 512 プロバイダー既定は 1,024 |
| 保持時間 | デフォルト 5 分、オプションで 1 時間 |
| 書き込みコスト | 1.25x (5m) / 2x (1h) |
思考
| ベンダー側パラメータ | thinking.type |
|---|---|
| 指定可能な値 | adaptive (default) · between_tools |
| デフォルト | adaptive, effort high リクエストで未指定の場合に適用 |
| 無効化の可否 | 非対応 |
| 思考の挙動 | Adaptive thinking is on by default. The lowest setting, between_tools, turns off up-front thinking and works at high effort or below; thinking {"type": "disabled"} and a manual {"type": "enabled", "budget_tokens": N} both return a 400 error. |
| パラメータ | reasoning_effort |
| 値 | minimal · low · medium · high ゲートウェイ側のパラメータ面——上のベンダーマッピングが適用されます |
モデル
| モダリティ | テキスト + 画像 → テキスト |
|---|
- Same price as Claude Sonnet 5
- 1M context at standard pricing with no long-context tier
- prompt-cache reads cost 0.1x input ($0.20/M)
- minimum cacheable prompt 512 tokens
- setting temperature, top_p or top_k to a non-default value returns a 400
- forced tool use returns an error
- text between tool calls comes back in thinking blocks
- up to 300k output tokens on the Message Batches API with the output-300k-2026-03-24 beta header
1つのプロンプト — ゲートウェイ経由で測定
Claude Sonnet 5.5 パス · 3 sentences
Enabling reasoning makes the model "think" before it answers, and that internal thinking is generated as extra text (reasoning tokens) that we are billed for at the same rate as output tokens, even though users never see it. Output tokens also cost several times more than input tokens, so a single request that used to produce a short answer can now generate many times more billable text behind the scenes. Because our request volume and prompts didn't change, the doubling comes almost entirely from this hidden thinking, and we can bring it down by limiting reasoning effort, turning it off for simple tasks, and reserving it for the workloads where the accuracy gain justifies the cost.
出力 205 tok レイテンシ 3.6 s
指示追従(ちょうど三文 — 数えられる)、読み手に合わせた調整(CFO 向けの語り口)、そして下のトークン計から見える隠れた思考の課金ギャップ。
Claude Sonnet 5.5 パス · 8/8 cases
**The bug:** `cur` counts the number of *adjacent equal pairs* rather than the number of items in the run. A run of length k contains only k−1 such pairs, so every result is one too small. The counter also starts at 0 and the loop starts at index 1, so a single-element list never updates `best` and returns 0 instead of 1. (An empty list should return 0.) In addition, `cur` resets to 0 on a mismatch, when it should reset to 1, because the item at `xs[i]` begins a new run of length 1. The claim that it "misses runs ending at the last index" is really the same off-by-one: the final run is counted, but as one less than its true length. Because `best` is updated inside the loop on every extension, the last run isn't skipped. […]
出力 444 tok レイテンシ 6.7 s
修正が実際に正しいか(実行可能か)、説明の密度、そして範囲の限られたタスクでのトークン効率。
Claude Sonnet 5.5 パス · 5/5 fields, guidance null
Here is the extracted JSON: ```json { "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": 412, "growth_pct": 11, "guidance": null } ``` Notes on the choices: - **quarter**: The text says "September quarter" without specifying a fiscal year or Q-number, so I kept the wording as stated. - **revenue_usd_m**: The text gives "412" for the topline, with no currency or unit stated explicitly. I've taken it as 412 million USD per the field definition, but the text does not confirm that, and the figure is constant-currency. - **growth_pct**: "Eleven and change" means slightly above 11, so 11 is an approximation. The exact figure isn't given. - **guidance**: Guidance is withheld pending the Q4 close, so no guidance is stated, hence null.
出力 300 tok レイテンシ 3.5 s
スキーマ遵守(フィールドをでっち上げない)、ハルシネーション圧(guidance は明示的に保留されている)、そして構造化出力の経路の違い。
Claude Sonnet 5.5 パス · 122 words, 0 banned, 1 question
**Introducing Cross-Provider Prompt Caching** Your API gateway now caches prompts across every connected LLM provider, whether you route traffic to OpenAI, Anthropic, Google, or a self-hosted model. Here's how it works: the gateway fingerprints each prompt prefix, stores it centrally, and applies the right provider-specific caching mechanism automatically. When you switch providers or fail over during an outage, your cached context follows you. Early testing shows input token costs dropping by up to 70% on repeated system prompts and long documents, with latency improvements of 40% or more on cache hits. Tired of paying to send the same instructions thousands of times a day? Enable it with one configuration flag. Cache rules, TTLs, and per-team isolation are available in the dashboard today.
出力 266 tok レイテンシ 3.8 s
制約の遵守(語数の上限、禁止語リスト、唯一の疑問文)、文体の指紋、そして長さの制御。
30 秒で Claude Sonnet 5.5 を使う
OpenAI 互換。base_url を差し替えるだけで、SDK はそのまま。POST /v1/chat/completions
from openai import OpenAI
client = OpenAI(
base_url="https://synthorai.io/v1",
api_key="sk-syn-...",
)
resp = client.chat.completions.create(
model="claude-sonnet-5-5",
messages=[{"role": "user", "content": "Summarize this diff"}],
reasoning_effort="medium",
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://synthorai.io/v1",
apiKey: "sk-syn-...",
});
const resp = await client.chat.completions.create({
model: "claude-sonnet-5-5",
messages: [{ role: "user", content: "Summarize this diff" }],
reasoning_effort: "medium",
});
console.log(resp.choices[0].message.content);curl https://synthorai.io/v1/chat/completions \
-H "Authorization: Bearer sk-syn-..." \
-H "Content-Type: application/json" \
-d '{
"model": "claude-sonnet-5-5",
"messages": [{"role": "user", "content": "Hello"}],
"reasoning_effort": "medium"
}'package main
import (
"context"
"fmt"
"github.com/openai/openai-go/v3"
"github.com/openai/openai-go/v3/option"
)
func main() {
client := openai.NewClient(
option.WithBaseURL("https://synthorai.io/v1"),
option.WithAPIKey("sk-syn-..."),
)
resp, _ := client.Chat.Completions.New(context.TODO(), openai.ChatCompletionNewParams{
Model: "claude-sonnet-5-5",
Messages: []openai.ChatCompletionMessageParamUnion{
openai.UserMessage("Summarize this diff"),
},
ReasoningEffort: openai.ReasoningEffortMedium,
})
fmt.Println(resp.Choices[0].Message.Content)
}import com.openai.client.OpenAIClient;
import com.openai.client.okhttp.OpenAIOkHttpClient;
import com.openai.models.chat.completions.*;
import com.openai.models.ReasoningEffort;
OpenAIClient client = OpenAIOkHttpClient.builder()
.baseUrl("https://synthorai.io/v1")
.apiKey("sk-syn-...")
.build();
ChatCompletion resp = client.chat().completions().create(
ChatCompletionCreateParams.builder()
.model("claude-sonnet-5-5")
.addUserMessage("Summarize this diff")
.reasoningEffort(ReasoningEffort.MEDIUM)
.build());
System.out.println(resp.choices().get(0).message().content().orElse(""));Claude Sonnet 5.5 について
- 100 万入力トークンあたり 2 ドル、100 万出力トークンあたり 10 ドルで、プロンプトキャッシュの読み取りは入力価格の 10%(100 万トークンあたり 0.20 ドル)、5 分キャッシュの書き込みは 2.50 ドル、1 時間キャッシュの書き込みは 4 ドルです。
- 長文コンテキストの割増なしで 1M トークンのコンテキストウィンドウと 128K の最大出力を備え、テキストと画像の入力を受け付けます。
- 適応型思考は既定で有効で、既定の強度は high です。
- 最も低い設定の between_tools は事前の思考をオフにし、high 以下で動作します。
- 一方、thinking を disabled にしたり budget_tokens を手動指定したりすると 400 エラーになります。
- temperature、top_p、top_k を既定以外の値にしても 400 が返り、ツールの強制使用はサポートされません。
- Anthropic は Claude Sonnet 5 で動いているコードに影響する 5 つの破壊的変更を挙げており、ツール呼び出しの間のテキストは thinking ブロックとして返るようになりました。
- キャッシュ可能な最小プロンプトは 512 トークンです。
- Synthorai は他のモデルと同じ OpenAI 互換 API で Claude Sonnet 5.5 を提供します。
よくある質問
Claude Sonnet 5.5 API は無料で試せますか?
はい。新規アカウントには 10 回のトライアル呼び出しと最大 $1 の無料クレジットが付与され、カード登録は不要です。入力 $2/M で計算すると、このクレジットだけで Claude Sonnet 5.5 に対して約 62 回の ~8K トークンのリクエストを送れます。
Claude Sonnet 5.5 は何が得意ですか?
ラインアップで速度と知性のバランスが最も良いモデル、Sonnet 5 と同価格:100 万トークンあたり入力 2 ドル・出力 10 ドル、between_tools で事前の思考をオフ。全体像はベンダー公式のリリースノートに基づく「このモデルについて」セクションをご覧ください。
Claude Sonnet 5.5 の料金はいくらですか?
Synthorai 上の Claude Sonnet 5.5 は入力 100 万トークンあたり $2、出力 100 万トークンあたり $10 です。ベンダー定価のままで、プラットフォーム手数料はありません。キャッシュ済み入力トークンは $0.2/M で課金されます。
Claude Sonnet 5.5 はプロンプトキャッシュに対応していますか?
はい。オプトイン方式で、安定したプレフィックスを cache_control ブレークポイントでマークします。キャッシュ済み入力トークンは $0.2/M(未キャッシュは $2/M)で課金されます。なお、キャッシュには 512 トークン以上の安定したプレフィックスが必要です(TTL デフォルト 5 分、オプションで 1 時間)。 プロンプトキャッシュガイド →
Claude Sonnet 5.5 を利用するには?
お使いの OpenAI SDK の base_url を "https://synthorai.io/v1" に向け、model="claude-sonnet-5-5" を設定すれば完了です。API キー 1 本でゲートウェイ上のすべてのモデルを利用できます。
Claude Sonnet 5.5 の知識カットオフはいつですか?
ベンダー公式ドキュメントによると、Claude Sonnet 5.5 の知識カットオフは 2026-06 です(2026-09-29 時点)。
関連モデル
比較
このページの値はすべてベンダー自身のドキュメント(上部にリンク)から転記し、確認した日付を付しています。価格はカタログ全体で比較しますが、ベンダーごとに定義が異なる仕様値は差異を明記するにとどめ、図表で比較はしません。当社が測定した数値はなく、スコアも付けていません。