---
format: "aidr-story-markdown/v1"
id: "729a527254ccbfde46350cbe648dc9c72fd14ccf989c9342f2acdc7c165a3917"
canonical_url: "https://aidr.today/729a5272?lang=en"
title: "OpenAI Opens GPT-Live-1 Voice API at $0.05 Per Minute to Cut Latency to 0.798 Seconds"
lang: "en"
requested_lang: "en"
available_langs: ["en","vi"]
translation_fallback: null
fallback_fields: []
published_at: "2026-09-10T18:55:19.000Z"
category: "Products"
topics: ["openai","gpt","voice-api","inference","agent","multimodal"]
source_urls: ["https://huggingnews.com/ai/openai-opens-gpt-live-1-voice-api-at-005-per-minute-to-cut-latency-to-07-d59c97fd","https://x.com/OpenAIDevs/status/2098099269551149398","https://x.com/OpenAIDevs/status/2098099372999446590","https://x.com/OpenAIDevs/status/2098099324509151532","https://x.com/juberti/status/2098102270332444922","https://x.com/thsottiaux/status/2098105060186374280","https://x.com/wallstengine/status/2098110480216883490","https://x.com/OpenAIDevs/status/2098109685366915316"]
summary: "Developers can now integrate a full duplex voice system into applications that distinguishes speech from background noise and supports natural interruptions. The model combines listening and speaking in a single architecture to eliminate handoffs between separate audio and text processes. Tool calling for spoken requests has an 87% pass rate, allowing AI agents to delegate complex reasoning to backend models such as GPT-6 Astra. The new API costs $0.05 per minute for the voice layer and lowers turn taking latency to 0.798 seconds from 1.41 seconds for GPT-Realtime-2.1. Based on the technology provided to 1 billion ChatGPT users, the system achieves an 86.2% score on Tau3 voice tasks compared to 45.7% for the previous version. Early deployments include restaurant service tools from Yelp Host and Hatch, and language lessons from Speak, which saw nearly 80% fewer interruptions in early testing."
---

# OpenAI Opens GPT\-Live\-1 Voice API at $0\.05 Per Minute to Cut Latency to 0\.798 Seconds

> [Open the canonical story](<https://aidr.today/729a5272?lang=en>)

**Published:** 2026-09-10T18:55:19.000Z
**Category:** Products
**Topics:** openai, gpt, voice\-api, inference, agent, multimodal

## Summary

Developers can now integrate a full duplex voice system into applications that distinguishes speech from background noise and supports natural interruptions\. The model combines listening and speaking in a single architecture to eliminate handoffs between separate audio and text processes\. Tool calling for spoken requests has an 87% pass rate, allowing AI agents to delegate complex reasoning to backend models such as GPT\-6 Astra\. The new API costs $0\.05 per minute for the voice layer and lowers turn taking latency to 0\.798 seconds from 1\.41 seconds for GPT\-Realtime\-2\.1\. Based on the technology provided to 1 billion ChatGPT users, the system achieves an 86\.2% score on Tau3 voice tasks compared to 45\.7% for the previous version\. Early deployments include restaurant service tools from Yelp Host and Hatch, and language lessons from Speak, which saw nearly 80% fewer interruptions in early testing\.

## Sources

- [Story source](<https://huggingnews.com/ai/openai-opens-gpt-live-1-voice-api-at-005-per-minute-to-cut-latency-to-07-d59c97fd>)
- [Story source](<https://x.com/OpenAIDevs/status/2098099269551149398>)
- [Supporting source](<https://x.com/OpenAIDevs/status/2098099372999446590>)
- [Supporting source](<https://x.com/OpenAIDevs/status/2098099324509151532>)
- [Supporting source](<https://x.com/juberti/status/2098102270332444922>)
- [Supporting source](<https://x.com/thsottiaux/status/2098105060186374280>)
- [Supporting source](<https://x.com/wallstengine/status/2098110480216883490>)
- [Supporting source](<https://x.com/OpenAIDevs/status/2098109685366915316>)

