Posts tagged cost
6 posts about cost.
-
GPT-5.6 Cost Guide: Prompt Caching 90% Off, Reasoning Effort
GPT-5.6's two cost levers, measured: explicit breakpoints bill cached input at 10% of the rate, and not sending reasoning_effort bills 1.5x as much as none.
-
Claude Fable 5 for Agents: Tool-Call Refusals, Cost vs GLM 5.2
Claude Fable 5 across five agent workloads vs glm-5.2, opus-4-8 and sonnet-5: mid-tool-call refusals, adaptive thinking, and cost that shifts 5-15x by shape.
-
Claude Sonnet 5's New Tokenizer: 41% More Tokens per Prompt
Claude Sonnet 5's new tokenizer makes the same text about 41% more tokens than Sonnet 4.6, reshaping cost, budgets, and cache eligibility on the gateway.
-
Transcription API Cost: 7 Models on the Same Audio
Seven transcription models, one multilingual audio set, one gateway: per-minute cost spans $0.0020 to $0.0164 and accuracy is not the differentiator.
-
GLM 5.2 Reasoning Effort: the Setting That Cuts Cost 20x (Measured)
Same coding answer: $0.0031 with reasoning effort set right vs $0.062 on GLM 5.2's unbounded default. 20x cheaper, 30x faster. How to set the dial per task.
-
Image Generation API Cost: 5 Models Compared ($0.006-$0.039)
Five image models, same prompts, one gateway: $0.006 to $0.039 per image at defaults, plus a quality knob that swings one model's bill 36x. Measured.