New Sign up free, 10 calls on us. Up to $1, no card needed.

Claude Fable 5 vs MiniMax M3

Neither Claude Fable 5 nor MiniMax M3 is in our catalogue any longer. The figures below are their last published rates; calls to either are no longer served.

vs

Which one, when

Both models share a 1000000-token context window and accept text and image input, so the real split is price and shape: minimax-m3 charges $0.3 input and $1.2 output against $10 and $50 for claude-fable-5, roughly 33x and 42x cheaper, and it also takes video input, allows up to 524288 output tokens, and lets thinking be switched off. Pick claude-fable-5 when you want Anthropic's always-on reasoning mode for chat, code and tool use and the higher rate is acceptable; pick minimax-m3 for high-volume, long-context or video work where cost and toggleable thinking matter more.

Benchmarks

MiniMax M3: the vendor has not published benchmark scores.

Above averageNo peer higherClaude Fable 586 / 9740 / 97
Claude Fable 5 MiniMax M3 other models measured peer average ★ no peer scored higher
SWE-Bench Pro
no peer scored higher 80.3%
N/A
BioMysteryBench hard
46.5%
N/A
OSWorld-Verified
no peer scored higher 85%
N/A
Cybergym
83.1%
N/A
HealthBench Professional
60.9%
N/A
Finance Agent v2
56.3%
N/A
Harvey Lab-AA
93.6%
N/A
GPQA Diamond
92.6%
N/A
Blueprint-Bench 2
38.6%
N/A
BrowseComp
87.4%
N/A
CharXiv (RQ) reasoning
no peer scored higher 88.9%
N/A

Vendor-published: Alibaba (Qwen) Anthropic ByteDance DeepSeek Google MiniMax Moonshot OpenAI Tencent Z.ai

Pricing

Claude Fable 5 MiniMax M3 Δ
Input / 1M tokens $10 $0.3 33×
Output / 1M tokens $50 $1.2 42×
Cache read / 1M tokens $1 $0.06 17×
Cache write 1.25x (5m) / 2x (1h) no separate charge -

Rates from the live catalogue at build time; each model page carries the current rate card.

Where they sit · input price per 1M tokens across all 76 chat models on this billing unit (log scale)

Claude Fable 5 · $10 MiniMax M3 · $0.3
$0.05 · Qwen3 VL Flash $30 · GPT-5.4 Pro

Capabilities

Claude Fable 5 MiniMax M3
Tool calling yes yes
Thinking control always on configurable
Structured output yes -
Prompt caching explicit (you mark the prefix) implicit (automatic)
Cache lifetime 5m default, 1h option not published
Minimum cached prefix 1024 tokens 512 tokens

Specs

Claude Fable 5 MiniMax M3
Input modalities text image text image video
Output modalities text text
Released 2026-06-09 2026-06-01
Knowledge cutoff 2026-01 -
Context window 1M 1M
Max output 128K 524K
Thinking parameter output_config.effort (thinking.type is adaptive-only and needs no configuration)
  • thinking.type
  • reasoning_split
Accepted values
effort
  • low
  • medium
  • high
  • xhigh
  • max

both "enabled" and "disabled" return 400

thinking.type
  • adaptive
  • disabled
reasoning_split
  • boolean
Default

thinking always on (adaptive)

effort
  • high
adaptive: thinking on, with the model deciding when extra reasoning helps

Specs are transcribed from each vendor’s documentation; a row a vendor does not publish is left out rather than inferred. Full sources: Claude Fable 5 · MiniMax M3

Switch between them with one line

Both ids are in every tab below; the highlighted pair of lines is the only edit. Same endpoint, same key, same request shape.

from openai import OpenAI

client = OpenAI(
    base_url="https://synthorai.io/v1",
    api_key="sk-syn-...",
)

resp = client.chat.completions.create(
    model="claude-fable-5",
    # model="minimax-m3",  # uncomment this line, comment the one above
    messages=[{"role": "user", "content": "Summarize this diff"}],
    reasoning_effort="medium",
)
print(resp.choices[0].message.content)

Get your API key →

FAQ

Which is cheaper, Claude Fable 5 or MiniMax M3?

MiniMax M3 is cheaper on the "Input / 1M tokens" row ($0.3 vs $10, 33× apart). Other rows may point the other way; the table above carries the full rate card, and real cost depends on your mix.

Can I A/B test Claude Fable 5 against MiniMax M3 without two integrations?

Yes. Both are served through the same OpenAI-compatible endpoint with one API key. Switching is a one-line change to the model id, so you can route a fraction of traffic to each and compare bills directly.

Do Claude Fable 5 and MiniMax M3 support prompt caching?

Yes. Both bill cache reads below their input rate, so warm-prefix workloads cost less than the list rates suggest. The exact cache-read rows are in the pricing table above.

Related comparisons

From our measured studies