OpenAI previewed Ultrafast mode: GPT-5.6 Sol at up to 14x speed, powered by Cerebras, generating up to 750 tokens per second. It launches first in the OpenAI API for a select customer group, then expands as capacity grows. Target workflows include realtime voice and support, commerce, coding and design, financial research, and security response—frontier intelligence where every second counts.

Key Takeaways

  • Frontier coding quality at near-instant completion speed is now a product mode, not just a benchmark demo.
  • 750 tokens/sec is Cerebras-backed; capacity, not the model card, is the rollout bottleneck.
  • Access is invite/select-customer preview, so production adoption still waits on fleet growth.
ADSponsored