🎁 New Sign up free, 10 calls on us. Up to $1, no card needed.
Live price comparison

Every model. Real prices, zero markup.

Compare transparent pay-as-you-go token prices across 9 providers · 100 models, USD per 1M tokens. Live from /api/pricing.

These are official list prices. Logged-in customers may see effective prices including workspace discounts on /console/pricing.

OpenAI

25 models
GPT family · frontier reasoning
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
16K
$1.25
$5.00
16K
$2.50
$10.00
16K
$2.50
$2.50
400K
$1.75
$14.00
$0.875
400K
$1.75
$14.00
$0.875
922K
$2.50
$15.00
$1.25
400K
$0.750
$4.50
$0.375
400K
$0.200
$1.25
$0.100
922K
$30.00
$180.00
$15.00
922K
$5.00
$30.00
$0.500
1.1M
$30.00
$180.00
1.1M
$5.00
$30.00
$0.500
1.1M
$1.00
$6.00
$0.100
1.1M
$5.00
$30.00
$0.500
1.1M
$2.50
$15.00
$0.250
$5.00
$40.00
$2.00
$8.00
$5.00
$32.00
$5.00
$30.00
$4.00
$16.00
$0.400
$4.00
$24.00
$0.400
$0.600
$2.40
$0.060
$0.10/s
$0.30/s
$0.006/min

Anthropic

10 models
Claude · long context, reliable tools
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
200K
$10.00
$50.00
$1.00
200K
$1.00
$5.00
$0.100
200K
$5.00
$25.00
$0.500
1M
$5.00
$25.00
$0.500
1M
$5.00
$25.00
$0.500
1M
$5.00
$25.00
$0.500
1M
$5.00
$25.00
$0.500
200K
$3.00
$15.00
$0.300
1M
$3.00
$15.00
$0.300
1M
$2.00
$10.00
$0.200

Google

22 models
Gemini · multimodal, large context
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
$0.016/min
$0.016/min
1.0M
$0.300
$2.50
$0.030
$0.300
$30.00
1.0M
$0.100
$0.400
$0.010
$0.500
$2.00
1.0M
$1.25
$10.00
$0.125
1.0M
$0.500
$3.00
$0.050
$2.00
$120.00
$0.500
$60.00
$0.250
$30.00
1.0M
$0.250
$1.50
$0.750
$4.50
1.0M
$2.00
$12.00
$0.200
1.0M
$1.50
$9.00
$0.150
1.0M
$0.300
$2.50
$0.030
1.0M
$1.50
$7.50
$0.150
$0.50/s
$0.15/s
$0.75/s
$0.15/s
$0.40/s

Alibaba

19 models
Qwen · cost-efficient, open weights
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
$0.002/min
$0.002/min
$0.002/min
$0.002/min
$0.002/min
$0.04/call
$0.07/call
$0.002/min
$0.002/min
256K
$1.20+
$6.00
$0.359
256K
$0.050+
$0.400
$0.022
256K
$0.200+
$1.60
$0.143
1.0M
$0.100
$0.400
$0.029
1.0M
$0.400+
$2.40
$0.115
256K
$0.250
$1.50
$0.050
1M
$2.50
$7.50
$0.500
1M
$0.400
$1.60
$0.080
$0.03/call
$0.07/call

DeepSeek

2 models
DeepSeek · strong reasoning, open weights
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
1M
$0.138
$0.275
$0.003
131K
$1.61
$3.22
$0.013

Moonshot

3 models
Kimi · long context
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
262K
$0.574
$3.01
$0.115
256K
$0.950
$4.00
$0.190
1M
$3.00
$15.00
$0.300

MiniMax

1 model
MiniMax · multimodal, agentic
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
1M
$0.300
$1.20
$0.060

Z.ai

4 models
GLM · bilingual, tool use
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
200K
$1.00
$3.20
$0.200
205K
$1.20
$4.00
$0.240
200K
$1.40
$4.40
$0.260
1.0M
$1.40
$4.40
$0.260

ByteDance

13 models
Doubao · high-throughput inference
Model · Capabilities
Context
Input/1M
Output/1M
Cache/1M
262K
$0.250+
$2.00
$0.050
262K
$0.250+
$2.00
$0.050
262K
$0.100+
$0.400
$0.020
262K
$0.500+
$3.00
$0.100
$7.00
$0.002/min
$1.00
$2.40
$0.03/call
$0.04/call
$0.04/call

Pricing notes

  • All prices are in USD. Billing is based on actual token usage.
  • Input tokens = your prompt. Output tokens = the model’s response.
  • Tiered pricing: a request’s entire token count is charged at the rate of the tier its input falls into. Larger inputs cross over to a higher tier.
  • Cache hit: when the same prompt prefix is reused (context caching), input tokens are billed at the discounted Cache/1M rate shown above.
  • Image and audio generation models are billed per API call.
  • Volume discounts available for enterprise customers. Contact us.

Found your model? Start calling it in 60 seconds.

Get your API key, free

Frequently asked questions

How do I compare LLM API prices across providers?

Synthorai's models page is a live price comparison across every major provider — sort by input or output price per million tokens, filter by provider, always in sync with the gateway's actual list prices. Every figure is exactly what a request draws from your balance, with a 0% platform surcharge.

Which LLM API providers does Synthorai support?

Synthorai is a multimodal AI gateway with one API and one bill for GPT (OpenAI), Claude (Anthropic), Gemini (Google), Qwen (Alibaba), DeepSeek, and more — across chat, image, and audio, with native prompt caching, BYOK, and zero data retention by default. The table above lists every model with its live price.

Which LLM API is the cheapest?

The table above is sorted by list price, so you can pick the cheapest model that meets your quality bar. Prompt caching bills cache hits at about 10% of the input price, and there is no platform fee — so effective cost is often lower than an aggregator at the same list price.