New Sign up free, 10 calls on us. Up to $1, no card needed.
Live price comparison

Every model. Real prices, zero markup.

Compare transparent pay-as-you-go token prices across 11 providers · 117 models, USD per 1M tokens. Live from /api/pricing.

These are official list prices. Logged-in customers may see effective prices including workspace discounts on /console/pricing.

Sort
Provider

OpenAI

28 models
GPT family · frontier reasoning
transcription
Input/1M
Output/1M
Cache/1M
$6.00Audio
$2.50
$10.00
-
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$1.75
$14.00
$0.875
chatcodetoolsreasoning
Input/1M
Output/1M
Cache/1M
$1.75
$14.00
$0.875
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$2.50
$15.00
$1.25
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$0.750
$4.50
$0.375
chatcodetoolsreasoning
Input/1M
Output/1M
Cache/1M
$0.200
$1.25
$0.100
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$30.00
$180.00
$15.00
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$5.00
$30.00
$0.500
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$30.00
$180.00
-
gpt-5.61.1M
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$5.00+
$30.00
$0.500
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$1.00
$6.00
$0.100
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$5.00+
$30.00
$0.500
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$2.50+
$15.00
$0.250
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$10.00+
$50.00
$1.00
image
Input/1M
Output/1M
Cache/1M
$5.00
$40.00
-
image
Input/1M
Output/1M
Cache/1M
$5.00
$30.00
-
realtimeaudio
Input/1M
Output/1M
Cache/1M
$32.00Audio
$4.00
$16.00
$0.400
realtimeaudio
Input/1M
Output/1M
Cache/1M
$32.00Audio
$4.00
$24.00
$0.400
video
Input/1M
Output/1M
Cache/1M
$0.10/s
-
-
video
Input/1M
Output/1M
Cache/1M
$0.30 - $0.50/s
-
-
speech
Input/1M
Output/1M
Cache/1M
$15.00
-
-
speech
Input/1M
Output/1M
Cache/1M
$30.00
-
-
transcription
Input/1M
Output/1M
Cache/1M
$0.006/min
-
-

Anthropic

10 models
Claude · long context, reliable tools
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$10.00
$50.00
$0.250
chatcodetools
Input/1M
Output/1M
Cache/1M
$1.00
$5.00
$0.100
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$5.00
$25.00
$0.500
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$5.00
$25.00
$0.500
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$5.00
$25.00
$0.500
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$5.00
$25.00
$0.500
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$3.00
$15.00
$0.300
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$3.00
$15.00
$0.300
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$3.00
$15.00
$0.300
chatcodethinkingtools
Input/1M
Output/1M
Cache/1M
$2.00
$10.00
$0.200

Google

24 models
Gemini · multimodal, large context
transcription
Input/1M
Output/1M
Cache/1M
$0.016/min
-
-
transcription
Input/1M
Output/1M
Cache/1M
$0.016/min
-
-
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$1.00Audio
$0.300
$2.50
$0.030
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$0.300Audio
$0.100
$0.400
$0.010
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$1.25Audio
$1.25
$10.00
$0.125
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$5.00Audio
$1.50
$9.00
$0.150
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$5.00Audio
$1.50
$7.50
$0.150
chatvisioncodetools
Input/1M
Output/1M
Cache/1M
$0.750Audio
$0.750
$3.75
$0.075

Alibaba

23 models
Qwen · cost-efficient, open weights
transcription
Input/1M
Output/1M
Cache/1M
$0.002/min
-
-
transcription
Input/1M
Output/1M
Cache/1M
$0.002/min
-
-
transcription
Input/1M
Output/1M
Cache/1M
$0.002/min
-
-
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.110
$0.800
$0.022
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.20+
$6.00
$0.359
chatvisionreasoningtools
Input/1M
Output/1M
Cache/1M
$0.050+
$0.400
$0.022
chatvisionreasoningtools
Input/1M
Output/1M
Cache/1M
$0.200+
$1.60
$0.143
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.100
$0.400
$0.029
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.400+
$2.40
$0.115
chatcodetools
Input/1M
Output/1M
Cache/1M
$0.250
$1.50
$0.050
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$2.50
$7.50
$0.500
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.400+
$1.60
$0.080
chatvisioncodereasoning
Input/1M
Output/1M
Cache/1M
$2.00
$6.00
$0.250
image
Input/1M
Output/1M
Cache/1M
$0.03/call
-
-

DeepSeek

5 models
DeepSeek · strong reasoning, open weights
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.280
$0.420
$0.028
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.32
$3.96
$0.132

Moonshot

3 models
Kimi · long context
chatvisioncodereasoning
Input/1M
Output/1M
Cache/1M
$0.574
$3.01
$0.115
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.950
$4.00
$0.190
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$3.00
$15.00
$0.300

MiniMax

2 models
MiniMax · multimodal, agentic
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.300
$1.20
$0.060
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.220
$0.900
$0.044

Z.ai

6 models
GLM · bilingual, tool use
glm-5200K
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.00
$3.20
$0.200
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.20
$4.00
$0.240
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.40
$4.40
$0.260
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.40
$4.40
$0.260
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$1.40
$4.40
$0.260
chatcodereasoningvision
Input/1M
Output/1M
Cache/1M
$0.150
$0.500
$0.030

ByteDance

14 models
Doubao · high-throughput inference
chatcodereasoningtools
Input/1M
Output/1M
Cache/1M
$0.500+
$3.00
$0.100
speech
Input/1M
Output/1M
Cache/1M
$30.00
-
-

Amazon

1 model
Nova · realtime speech
realtimeaudio
Input/1M
Output/1M
Cache/1M
$3.00Audio
$0.060
$0.240
-

Tencent

1 model
Hunyuan · long context, tool use
chatcodetoolsreasoning
Input/1M
Output/1M
Cache/1M
$0.140
$0.580
$0.035

Pricing notes

  • All prices are in USD. Billing is based on actual token usage.
  • Input tokens = your prompt. Output tokens = the model’s response.
  • Tiered pricing: a request’s entire token count is charged at the rate of the tier its input falls into. Larger inputs cross over to a higher tier.
  • Cache hit: when the same prompt prefix is reused (context caching), input tokens are billed at the discounted Cache/1M rate shown above.
  • Image and audio generation models are billed per API call.
  • Volume discounts available for enterprise customers. Contact us.

Found your model? Start calling it in 60 seconds.

Get your API key

Frequently asked questions

How do I compare LLM API prices across providers?

Synthorai's models page is a live price comparison across every major provider - sort by input or output price per million tokens, filter by provider, always in sync with the gateway's actual list prices. Every figure is exactly what a request draws from your balance, with a 0% platform surcharge.

Which LLM API providers does Synthorai support?

Synthorai is a multimodal AI gateway with one API and one bill for GPT (OpenAI), Claude (Anthropic), Gemini (Google), Qwen (Alibaba), DeepSeek, and more - across chat, image, and audio, with native prompt caching, BYOK, and zero data retention by default. The table above lists every model with its live price.

Which LLM API is the cheapest?

The table above is sorted by list price, so you can pick the cheapest model that meets your quality bar. Prompt caching bills cache hits at about 10% of the input price, and there is no platform fee - so effective cost is often lower than an aggregator at the same list price.

How do I compare speech-to-text API pricing?

Audio models bill two different ways - per minute of audio or per token - so sticker prices aren't directly comparable. Our speech-to-text API pricing comparison converts 14 models to the same per-minute unit, from $0.002 to $0.016 a minute.