DeepSeek introduced automatic API Prompt Caching across its developer platform. When re-sending system prompts, large repository contexts, or multi-turn chat history, cached input tokens are billed at just ¥0.1 per million tokens (~$0.014/M), reducing operating expenses for autonomous coding agents by up to 85%.
Key Takeaways
- ✓Cached input tokens cost only ¥0.1 (~$0.014) per million tokens with automated 64k prefix reuse.
- ✓Zero manual cache key configuration: gateway automatically handles prefix hash caching.
- ✓Dramatically lowers API costs for terminal agents (Aider, Cline, OpenHands) loading full repos.