DeepSeek opened V4-Flash developer preview with 1M context window and 180 tokens/second inference throughput at $0.02/1M.
Key Takeaways
- ✓180 tokens/sec generation speed (3x faster than baseline), tailored for sub-50ms Tab completions
- ✓1M context window processes full code repositories without chunking or losing semantic detail
- ✓Free public test API keys now available via DeepSeek Open Platform