Posts tagged caching
2 posts about caching.
-
Kimi K3 API Pricing, Measured: Turn Off the 'Always-On' Reasoning
Kimi K3's docs say reasoning can't be disabled. reasoning_effort:'none' works and cuts simple queries 6x. Measured: effort dial, cache floor, 9-language rates.
-
GPT Live API Pricing: gpt-realtime Speaking Costs 4x
GPT Live is the ChatGPT feature, not an API; behind it is gpt-realtime-2.1: $0.019/min listening, $0.077/min speaking, silence free, cached replay 1/80th.