SurviAGI
← All updates
OpenAI · 2026-09-10 · Benchmark result

OpenAI publishes GPT-Live-1 production voice agent benchmark results

OpenAI shared benchmark results for GPT-Live-1 covering task completion, multi-turn conversation and turn-taking, response latency, and tool use for production voice agents.

This update bears on 4 kinds of work in 4 markets. Highest level accepted: L2 Partial automation.

Highest accepted
L2Partial automation
Work
4
Markets
4
Views
54.5K

The work it bears on

Highest level first. Open a kind of work to see everything else that moved it.

Published at