← All updates
OpenAI · 2026-09-10 · Benchmark result
OpenAI publishes GPT-Live-1 production voice agent benchmark results
OpenAI shared benchmark results for GPT-Live-1 covering task completion, multi-turn conversation and turn-taking, response latency, and tool use for production voice agents.
This update bears on 4 kinds of work in 4 markets. Highest level accepted: L2 Partial automation.
- Highest accepted
- L2Partial automation
- Work
- 4
- Markets
- 4
- Views
- 54.5K
The work it bears on
Highest level first. Open a kind of work to see everything else that moved it.
- L2Partial automationL2 Partial automationClaimed L3 · Accepted L2From the publisher only
- L2Partial automationL2 Partial automationClaimed L3 · Accepted L2From the publisher only
- L2Partial automationL2 Partial automationFrom the publisher only
- L2Partial automationL2 Partial automationClaimed L3 · Accepted L2From the publisher only