DeepSeek-V4-Flash-Vision-Exp is live on the API: it matches V4-Flash on text—including agents, reasoning, and world knowledge—while jumping on multimodal agent benchmarks to near Claude Opus 4.8. A free Files API and DeepSeek Harness 0.1.1 ship with it; images bill at V4-Flash rates (up to 384 tokens each) and drop into existing agent frameworks.
Key Takeaways
- ✓The experimental vision model keeps V4-Flash text/agent strength while pushing multimodal agent scores close to Opus 4.8
- ✓Chat Completions, Messages, and Responses are supported; images mix in via base64, URLs, or the Files API
- ✓Files API is free with file_id reuse, and Harness 0.1.1 supports the model out of the box for agent stacks