OpenAI が Build more natural voice experiences with GPT‑Live‑1 in the API をリリース
OpenAI から Build more natural voice experiences with GPT‑Live‑1 in the API がリリースされました。
▸何が変わったのか
September 10, 2026
Product
Release
Build more natural voice experiences with GPT‑Live‑1 in the API
GPT‑Live‑1 brings ChatGPT’s natural, full-duplex conversations to the API, with more control over how voice agents speak and act.
Loading…
Share
We’re launching GPT‑Live‑1 in the API, giving developers a powerful, natural voice model for building voice-enabled apps and business workflows.
First introduced in ChatGPT
, GPT‑Live‑1 is capable of listening and speaking at the same time, and, as
seen with Codex and ChatGPT Work
(opens in a new window)
, can delegate deeper reasoning and actions to the models and tools it is paired with.
For the API release of GPT‑Live‑1, we’ve focused on new capabilities that let developers steer and customize voice experiences around their users, workflows, and goals. A core GPT‑Live‑1 strength, smooth interruption handling, is already delivering business impact: in early evaluations, Speak found that GPT‑Live‑1 gave learners more time to think before the language tutor responded, cutting interruptions by almost 80% versus previous turn-based systems.
Key strengths of GPT‑Live‑1 in the API:
Interruption handling:
Improves interruption handling via a single model that reasons over incoming and outgoing audio together, avoiding the latency and brittle handoffs of chained STT–LLM–TTS architectures.
Reasoning & tool calling delegation:
GPT‑Live‑1 can delegate reasoning and tool calls to a backend text model like GPT‑6 Astra or a third-party model.
Tone, pace, and style:
Lets developers shape an agent’s tone, pace, and conversational style through the system prompt.
Silent context management & background noise:
Better handles background noise and silence without interrupting the conversation or narrating every step out loud.
Long-session reliability:
Improves context retention and conversational quality across extended interactions.
Telephony support:
Enables deployment of full-duplex voice agents for phone calls, from restaurant reservations to customer support.
Try GPT-Live-1
Start a session and speak naturally. Interrupt, laugh, change your mind – try it at home or in a loud space like a coffee shop or city street.
Start session
See what it can do
Talk over it—naturally.
Ask for help, then interrupt mid-response to change the question or add detail.
Take it with you.
Try a conversation while walking outside or with everyday background noise, and see how it stays with you.
Make it playful.
Laugh, hesitate, use short acknowledgments, or briefly talk to someone nearby—then continue the conversation.
This demo is time-limited. By using it, you agree to OpenAI’s
Terms
and acknowledge our
Privacy Policy
.
Simplify your voice-agent architecture and reduce voice latency
Traditional voice agents stitch together speech-to-text, a reasoning model, and text-to-speech. Each handoff adds latency and creates more opportunities to lose timing, context, or the natural rhythm of a conversation. Developers are often the ones left coordina
▸Hacker Newsの反応
現時点ではまだコメントがゼロ。静観モードの人が多く、具体的な反応はこれから出てきそうだ。
▸Redditの反応
GPT-Live-1のAPI公開に対し、開発者たちは「わざわざ特定のプラットフォームに登録したくない」という利便性の不満や、「レストラン予約自動化なんてAGIの本来の目的じゃないだろ」という皮肉な反応が見られる。技術そのものへの興奮よりも、ビジネスモデルやユースケース選びへの冷めた視線が支配的だ。
「またしても自動レストラン予約のデモか。AGIもソフトウェアの存在意義も、結局この体験を自動化することにあるのかと呆れちゃう。ラボがこれしかできないのが不思議でならない。」
「OpenRouter経由で使えるようになってほしい。特定の提供元のプラットフォームに登録して、また一つ請求対象を増やしたくないから。」
「この機能のデモとして、最も使い道がなさそうな例を選んだな。」
SOURCE: OpenAI (2026-09-10)


