Per the OpenAI API changelog, the Decisions API entered public beta on 2026-10-06 with gpt-6-luna: a dedicated POST /v1/decisions endpoint takes text/images plus a list of questions and returns typed predicate, choice or score answers with probabilities, about 10x faster than the Responses API. Billing is input-only at $0.10 per 1M tokens (no output or cache charges); ZDR and HIPAA are supported and GA is expected within weeks.
Key Takeaways
- ✓Speed: typed answers about 10x faster than the Responses API (Decisions guide)
- ✓Price: input-only $0.10 per 1M tokens, no output or cache read/write fees; regular gpt-6-luna is $0.10 in / $0.50 out
- ✓3 question types: predicate (0–1 probability), choice (value + probabilities + confidence), score (probability-weighted ordinal level)
- ✓Compliance: ZDR and HIPAA eligible, US and Europe (EEA + Switzerland) data residency; GA expected in weeks
- ✓Limits: gpt-6-luna only; images must be inline base64, no hosted URLs or file_id

Key Decision Metrics at a Glance
Turn your technical choice into a development budget
Compare 40 dev plans & simulate token costs vs $20/mo subscriptions
Project Links & Resources
Direct AccessIn-Depth Technical Analysis
OpenAI moved the Decisions API to public beta on October 6, 2026 (API changelog: https://developers.openai.com/api/docs/changelog). It is a dedicated POST /v1/decisions endpoint, currently served only by gpt-6-luna, that evaluates shared text or text+image input against a list of developer-defined questions and returns an answers array. Three question types exist: predicate (probability a condition is true), choice (one of supplied values plus per-option probabilities and a confidence field) and score (probability-weighted average over ordered levels, indices from 0). OpenAI says it returns typed answers about 10x faster than the Responses API and positions it for classification, routing and an agent's next action, while Structured Outputs and function calling remain the tools for free-form objects and tool calls. Pricing is input-only at $0.10 per 1M tokens with no output, cache-read or cache-write charges; regional premiums and long-context multipliers apply. It supports Zero Data Retention and HIPAA for eligible customers and US/EU data residency. Images must be inline base64 data URLs. No accuracy benchmarks were published. A Playground is available, a voice integration guide covers Live API client delegation, and GA is expected in the coming weeks. The same changelog day also collapsed API usage tiers from five to three (Build, Launch, Grow).
Benchmark side-by-side against alternatives, or calculate monthly token cost vs subscription break-even.
Discussion & Comments
0Sign in to join the discussion
Connect with AI developers to exchange benchmark insights.