Agents & InferenceOpenAI

GPT-Live uses turnless speech model for continuous low-latency voice AI

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

GPT-Live cuts end-to-end voice latency to ~300ms, making real-time back-and-forth feel human. This lets you ship voice agents that don’t frustrate users with pauses, but it demands sub-100ms network hops and GPU colo near users to hit the latency budget.