新規 無料登録、10回の呼び出しを進呈。最大 $1、カード不要。

Claude Fable 5.1 vs Claude Sonnet 5.5

Claude Fable 5.1 は招待制で提供されています。下記の数値は実際のライブレートですが、呼び出しにはまずワークスペースの権限付与が必要となるため、この比較を基に開発を行う前にアクセス権限をご依頼ください。

vs

いつ、どちらを使うか

これらはClaudeのラインナップにおいて2ティア離れています:claude-fable-5-1はAnthropicが高度な推論と長期的なエージェントワーク向けと位置付ける最上位ティアであり、claude-sonnet-5-5はFableが遅いのに対して高速であると記載されている下位ティアです。どちらもテキストと画像を入力として受け付け、1000000トークンのコンテキストと128000の最大出力を備えているため、価格がその差を示しています:claude-fable-5-1は入力$10・出力$50であり、入力$2・出力$10であるclaude-sonnet-5-5の5xの価格で、キャッシュ読み取りは$0.2に対して$0.25です。日常的および大量のトラフィックはclaude-sonnet-5-5で実行し、それに見合う長期的なタスクにはclaude-fable-5-1へエスカレーションしてください。

ベンチマーク

Claude Sonnet 5.5:ベンダーはベンチマークのスコアを公表していません。

平均超え他モデル以上Claude Fable 5.119 / 213 / 21
Claude Fable 5.1 Claude Sonnet 5.5 測定された他モデル 測定対象の平均 ★ 他モデルに上回られていない
DeepSWE 1.1
67.4%
N/A
OSWorld 2.0 partial
80.7%
N/A
HealthBench Professional
58.1%
N/A
Terminal-Bench-Science 0.1
52.6%
N/A
GPQA Diamond
93.7%
N/A
AutomationBench
31.4%
N/A
Chartography with tools
88.4%
N/A

ベンダー公表: Alibaba (Qwen) Anthropic ByteDance DeepSeek Google MiniMax Moonshot OpenAI Tencent Z.ai

料金

Claude Fable 5.1 Claude Sonnet 5.5 Δ
入力 / 1Mトークン $10 $2 5×
出力 / 1Mトークン $50 $10 5×
キャッシュ読み取り / 1Mトークン $0.25 $0.2 1.3×
キャッシュ書き込み 1.25x (5m) / 2x (1h) 1.25x (5m) / 2x (1h) -

ビルド時のライブカタログの料金です。各モデルのページには現在の料金カードが記載されています。

位置付け — この課金単位におけるすべての76個のチャットモデル全体の1Mトークンあたりの入力料金 (対数スケール)

機能

Claude Fable 5.1 Claude Sonnet 5.5
ツール使用 あり あり
思考コントロール 常時オン 常時オン
構造化出力 あり あり
プロンプトキャッシング 明示的 (プレフィックスを自分で指定) 明示的 (プレフィックスを自分で指定)
キャッシュ有効期間 5m default, 1h option 5m default, 1h option
最小キャッシュプレフィックス 1024 トークン 1024 トークン

仕様

Claude Fable 5.1 Claude Sonnet 5.5
入力モダリティ テキスト 画像 テキスト 画像
出力モダリティ テキスト テキスト
リリース 2026-09-01 2026-09-28
知識のカットオフ 2026-06 2026-06
コンテキストウィンドウ 1M 1M
最大出力 128K 128K
思考パラメータ output_config.effort (thinking is adaptive-only and always on) thinking.type
指定可能な値
effort
  • low
  • medium
  • high
  • xhigh
  • max
thinking.type
  • adaptive (default)
  • between_tools
デフォルト

thinking always on (adaptive)

effort
  • high
adaptive, effort high

仕様は各ベンダーのドキュメントから転記されています。ベンダーが公開していない行は推測せず、省略しています。 すべての出典: Claude Fable 5.1 · Claude Sonnet 5.5

1つのプロンプト、両方のモデル — ゲートウェイ経由で測定

プロンプト Explain to a CFO, in exactly three sentences, why our LLM bill doubled after we enabled reasoning. チェック ちょうど 3 文

Claude Fable 5.1 パス · 3 sentences

When we enabled reasoning, the model started generating a long chain of internal "thinking" tokens before producing each answer, and we're billed for those tokens at the same output rate even though users never see them. In practice, that hidden reasoning often runs several times longer than the visible response, so the total tokens per request roughly doubled while the number of requests stayed flat. The upside is measurably better accuracy on complex tasks, so the right conversation is whether that quality gain justifies the cost—and if not, we can restrict reasoning to the workflows that actually need it rather than turning it on globally.

出力 186 tok レイテンシ 3.2 s

Claude Sonnet 5.5 パス · 3 sentences

Enabling reasoning makes the model "think" before it answers, and that internal thinking is generated as extra text (reasoning tokens) that we are billed for at the same rate as output tokens, even though users never see it. Output tokens also cost several times more than input tokens, so a single request that used to produce a short answer can now generate many times more billable text behind the scenes. Because our request volume and prompts didn't change, the doubling comes almost entirely from this hidden thinking, and we can bring it down by limiting reasoning effort, turning it off for simple tasks, and reserving it for the workloads where the accuracy gain justifies the cost.

出力 205 tok レイテンシ 3.6 s

指示追従(ちょうど三文 — 数えられる)、読み手に合わせた調整(CFO 向けの語り口)、そして下のトークン計から見える隠れた思考の課金ギャップ。

プロンプト This function is supposed to return the longest run of consecutive equal items, but callers report it is off by one on single-element inputs and misses runs that end at the last index. Fix it and explain the bug in one paragraph. チェック 修正がテストを通る

Claude Fable 5.1 パス · 8/8 cases

```python def longest_run(xs): if not xs: return 0 best = cur = 1 for i in range(1, len(xs)): cur = cur + 1 if xs[i] == xs[i - 1] else 1 best = max(best, cur) return best ``` The bug is that `cur` counts *matches between adjacent elements* rather than *elements in the run*: it starts at 0 and only increments when `xs[i] == xs[i-1]`, so a run of length k produces a count of k−1. That's why a single element returns 0 instead of 1, and why every other run also comes out one short — the "misses runs at the end" reports are the same off-by-one seen on inputs like `[1, 2, 2]`, where the trailing run of length 2 is reported as 1 (the original code does update `best` inside the loop, so end-of-list runs aren't actually skipped, just undercounted). […]

出力 882 tok (+507 思考) レイテンシ 3.4 s

Claude Sonnet 5.5 パス · 8/8 cases

**The bug:** `cur` counts the number of *adjacent equal pairs* rather than the number of items in the run. A run of length k contains only k−1 such pairs, so every result is one too small. The counter also starts at 0 and the loop starts at index 1, so a single-element list never updates `best` and returns 0 instead of 1. (An empty list should return 0.) In addition, `cur` resets to 0 on a mismatch, when it should reset to 1, because the item at `xs[i]` begins a new run of length 1. The claim that it "misses runs ending at the last index" is really the same off-by-one: the final run is counted, but as one less than its true length. Because `best` is updated inside the loop on every extension, the last run isn't skipped. […]

出力 444 tok レイテンシ 6.7 s

修正が実際に正しいか(実行可能か)、説明の密度、そして範囲の限られたタスクでのトークン効率。

プロンプト Extract a JSON object with fields {company, quarter, revenue_usd_m, growth_pct, guidance} from this text. Use null for anything not stated; add no other fields. チェック 妥当な JSON、スキーマ厳密

Claude Fable 5.1 パス · 5/5 fields, guidance null

```json { "company": "Meridian Robotics", "quarter": "Q3 (September quarter)", "revenue_usd_m": 412, "growth_pct": 11, "guidance": null } ``` Notes on judgment calls (not part of the object): - **revenue_usd_m**: The text says "412 for the topline" but never states the currency or that the figure is in millions; it's also on a constant-currency basis and prefaced with "call it," so it's approximate. If you want strict literalism, this should be `null`. - **growth_pct**: "Eleven and change" means slightly above 11%; the exact figure isn't given. - **quarter**: "September quarter" is Q3 on a calendar basis, consistent with the reference to a pending Q4 close. - **guidance**: Explicitly withheld, so `null`.

出力 633 tok (+349 思考) レイテンシ 3.8 s

Claude Sonnet 5.5 パス · 5/5 fields, guidance null

Here is the extracted JSON: ```json { "company": "Meridian Robotics", "quarter": "September quarter", "revenue_usd_m": 412, "growth_pct": 11, "guidance": null } ``` Notes on the choices: - **quarter**: The text says "September quarter" without specifying a fiscal year or Q-number, so I kept the wording as stated. - **revenue_usd_m**: The text gives "412" for the topline, with no currency or unit stated explicitly. I've taken it as 412 million USD per the field definition, but the text does not confirm that, and the figure is constant-currency. - **growth_pct**: "Eleven and change" means slightly above 11, so 11 is an approximation. The exact figure isn't given. - **guidance**: Guidance is withheld pending the Q4 close, so no guidance is stated, hence null.

出力 300 tok レイテンシ 3.5 s

スキーマ遵守(フィールドをでっち上げない)、ハルシネーション圧(guidance は明示的に保留されている)、そして構造化出力の経路の違い。

プロンプト Write a 120-word product announcement for an API gateway feature that caches prompts across providers. Forbidden words: "seamless", "unlock", "game-changer", "revolutionize", "empower". Exactly one sentence must be a question. チェック 120 語、禁止語 0

Claude Fable 5.1 パス · 121 words, 0 banned, 1 question

**Introducing Cross-Provider Prompt Caching for the Gateway.** Starting today, your API gateway can cache prompts once and reuse them across OpenAI, Anthropic, Google, and any other provider you route to. Long system instructions, retrieval context, and few-shot examples are stored at the gateway layer, so switching models no longer means paying to resend the same tokens. Why send the same 8,000-token preamble to three different vendors when you could send it once? Our cache normalizes prompt segments, fingerprints them, and serves matching prefixes from memory before the request ever reaches an upstream API. Early customers report token costs falling by 30 to 60 percent on high-volume workloads, with median latency dropping alongside. Available today on all Team and Enterprise plans.

出力 1722 tok (+1463 思考) レイテンシ 2.9 s

Claude Sonnet 5.5 パス · 122 words, 0 banned, 1 question

**Introducing Cross-Provider Prompt Caching** Your API gateway now caches prompts across every connected LLM provider, whether you route traffic to OpenAI, Anthropic, Google, or a self-hosted model. Here's how it works: the gateway fingerprints each prompt prefix, stores it centrally, and applies the right provider-specific caching mechanism automatically. When you switch providers or fail over during an outage, your cached context follows you. Early testing shows input token costs dropping by up to 70% on repeated system prompts and long documents, with latency improvements of 40% or more on cache hits. Tired of paying to send the same instructions thousands of times a day? Enable it with one configuration flag. Cache rules, TTLs, and per-team isolation are available in the dashboard today.

出力 266 tok レイテンシ 3.8 s

制約の遵守(語数の上限、禁止語リスト、唯一の疑問文)、文体の指紋、そして長さの制御。

1行で切り替え

以下のすべてのタブには両方のIDが含まれています — 変更箇所はハイライトされた2行のみです。エンドポイント、キー、リクエスト形式はすべて同じです。

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.chat.completions.create(
    model="claude-fable-5-1",
    # model="claude-sonnet-5-5",  # この行をアンコメントし、上の行をコメントアウトします
    messages=[{"role": "user", "content": "Summarize this diff"}],
    reasoning_effort="medium",
)
print(resp.choices[0].message.content)

API キーを取得 →

FAQ

Claude Fable 5.1 と Claude Sonnet 5.5 ではどちらが安いですか?

入力 / 1mトークン においては Claude Sonnet 5.5 の方が安価です($2 対 $10、5.0× の差)。他の行では逆になる可能性があります — 上の表には完全な情報が記載されており、実際のコストは組み合わせに依存します。

2つの統合を行わずに Claude Fable 5.1 と Claude Sonnet 5.5 のA/Bテストを実施できますか?

はい。両方とも1つのAPIキーで同じOpenAI互換エンドポイントを通じて提供されます — モデル文字列を1行変更するだけで切り替えられるため、トラフィックの一部をそれぞれにルーティングし、請求額を直接比較できます。

Claude Fable 5.1 と Claude Sonnet 5.5 はプロンプトキャッシングをサポートしていますか?

はい — どちらのモデルでもキャッシュ読み込みは入力レートよりも低く請求されるため、ウォームプレフィックスのワークロードは定価が示すよりも低コストになります。キャッシュ読み込みの正確な行は、上の料金表に記載されています。

関連する比較

当社の測定調査より