DeepSeek opened V4-Flash developer preview with 1M context window and 180 tokens/second inference throughput at $0.02/1M.

Key Takeaways

  • 180 tokens/sec generation speed (3x faster than baseline), tailored for sub-50ms Tab completions
  • 1M context window processes full code repositories without chunking or losing semantic detail
  • Free public test API keys now available via DeepSeek Open Platform