Introducing the Decisions API
OpenAI · 2026-10-06 · official · 117,864 views
What's in the video
Description written by Gemini, which watched and listened to the whole video.
Summary
Romain Huet (Head of Developer Experience at OpenAI) and Charlie Guo (Developer Experience at OpenAI) introduce OpenAI’s Decisions API. The API is powered by GPT-6 Luna and designed to return structured categorical decisions from text and visual inputs with sub-second latency.
What is shown
- Text extraction & lead routing UI: Demonstrating unstructured order text extraction into form fields [00:48] and inbound sales lead qualification [01:03] classifying company type, size, requirements, and next steps in 81 ms across six decisions [01:14].
- Vision lane-driving simulation: A top-down 2D driving game where camera frames of the road and obstacles are processed in real time by the Decisions API to select lane 1, 2, or 3 to steer the car clear of roadblocks [01:31].
- Voice avatar facial expressions: A conversational web application pairing GPT-Live 1 for voice with the Decisions API, which dynamically selects facial expressions for an animated frog character corresponding to conversational sentiment [02:17].
- Robotics object tracking with "Lavender": A tabletop programmable robot (a preview unit of Hugging Face’s Microdot robot) running GPT-Live and the Decisions API. The camera feed evaluates video frames to control head orientation to track a green apple [03:39], follow "the fruit" when asked ambiguously [04:12], and choose to track an Xbox controller over an apple when asked "what's more fun to play with?" [04:31].
- Post-credits blooper: Romain asks the robot if it is excited about the Decisions API, and the robot shakes its head horizontally [05:27].
Claims & numbers
- The Decisions API runs on GPT-6 Luna (presenter says at [00:19]).
- Focusing the model on constrained multiple choices makes it nearly 10 times faster than full generative output while maintaining image understanding, broad language support, and safety protection (presenter says at [00:28]).
- In the inbound sales demo, the server-side Decisions API processes and returns six decisions in 81 ms / less than 100 ms (presenter and UI show at [01:13]).
Notable quotes
- "By focusing the model on just a few choices, we can make it nearly 10 times faster..." — Romain Huet [00:28]
- "On the server side, the Decisions API here responded in less than 100 milliseconds, very impressive." — Romain Huet [01:11]
- "GPT-Live 1 handles the voice, while the Decisions API is choosing which expression the character should use as I talk." — Charlie Guo [02:20]
Assessment
This is an official OpenAI launch video demonstrating real-time interactive developer demos across web apps, vision navigation, animated agents, and hardware robotics. The displayed latencies (such as the 81 ms server processing time and real-time vision-based camera tracking) are live functional proof-of-concepts designed to highlight low-latency classification rather than benchmark evaluations.
Described by gemini-3.8-flash on 2026-10-07 from the video's audio and frames.