DeepSeek-V4-Flash-Vision-Exp is live on the API: it matches V4-Flash on text—including agents, reasoning, and world knowledge—while jumping on multimodal agent benchmarks to near Claude Opus 4.8. A free Files API and DeepSeek Harness 0.1.1 ship with it; images bill at V4-Flash rates (up to 384 tokens each) and drop into existing agent frameworks.

Key Takeaways

  • The experimental vision model keeps V4-Flash text/agent strength while pushing multimodal agent scores close to Opus 4.8
  • Chat Completions, Messages, and Responses are supported; images mix in via base64, URLs, or the Files API
  • Files API is free with file_id reuse, and Harness 0.1.1 supports the model out of the box for agent stacks
ADSponsored