DeepSeek introduced automatic API Prompt Caching across its developer platform. When re-sending system prompts, large repository contexts, or multi-turn chat history, cached input tokens are billed at just ¥0.1 per million tokens (~$0.014/M), reducing operating expenses for autonomous coding agents by up to 85%.

Key Takeaways

  • Cached input tokens cost only ¥0.1 (~$0.014) per million tokens with automated 64k prefix reuse.
  • Zero manual cache key configuration: gateway automatically handles prefix hash caching.
  • Dramatically lowers API costs for terminal agents (Aider, Cline, OpenHands) loading full repos.
ADSponsored