Posts tagged deepseek
2 posts about deepseek.
-
DeepSeek V4 Pro GA vs Preview, Measured: 18-62% Less Thinking
deepseek-v4-pro-0813 vs the preview build on identical tasks: 18-62% fewer reasoning tokens, a fixed 8,192-token runaway, and a lost 'I don't know'.
-
Open-Weight LLM Caching: Why Yours Is Provider Roulette
For open-weight LLMs, prompt caching is solved in the inference engine and broken by routing. A five-layer map, measured across DeepSeek, Qwen and Kimi.