O
OPENAI / gpt-live-1

GPT‑Live 1

Real-time voice model for expressive conversations, natural interruptions, and tool-enabled workflows.

01 / SUMMARY

The essential

ORIENTATIONReal-time voice conversation

Primary use stated or inferred cautiously from official documentation.

CONTEXTReal-time session

Maximum input capacity when published by the source.

EXITStreaming audio and text

Documented maximum generation limit.

ACCESSOpenAI API

Verified availability channels.

MODEL TYPEVoice and audio · Text and reasoning

Primary functional family and verified specialties.

LICENSEProprietary

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDgpt-live-1
Model typeVoice and audio · Text and reasoning
Access modelPaid proprietary
LicenseProprietary
DeploymentHosted API
Release2026
Knowledge cutoffNot published
EntranceText · Audio
ExitText · Audio
Context windowReal-time session
maximum outputStreaming audio and text
ReasoningNot applicable or not published
Published toolsFeatures · Voice sessions
Structured OutingsNot applicable or not published
Batch processingNot published
Prompt cacheNot published
Fine-tuningNot published
Verified platformsOpenAI Realtime API
BEST SUITED FOR

Voice agents, conversational support, and interactive spoken experiences.

WORTH MONITORING

Measure end-to-end latency, interruptions, cost per audio, and multilingual behavior.

03 / CAPABILITIES

What it can do

01

Orientation

Voice agents, conversational support, and interactive spoken experiences.

02

Context

Real-time session · Streaming audio and text

03

Tools and integration

Functions · Voice sessions

04

Access

OpenAI API

Input modalities
TextAudio
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard inputNot publishedStandard API rate
Cached inputNot publishedReading reused prefixes
Cache write or storageNot publishedThe condition varies by provider
Standard outputNot publishedMay include reasoning tokens
Batch inputNot publishedAsynchronous processing
Batch outputNot publishedAsynchronous processing

Consult the primary source: the billing unit depends on the model type and access channel.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

BenchmarkResultMetricSource
Evaluación de conversación GPT‑LiveMain modelNaturalness and interruptionsView source ↗

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

GPT‑Live 1

Version added to Inferama's verified catalog.

ANALYSIS

Related analysis

GPT‑Live Conversation Evaluation: How to Measure Turn-Taking, Interruptions, and Task Resolution Without Mistaking Smooth Dialogue for a Reliable Agent
ANALISIS

GPT‑Live Conversation Evaluation: How to Measure Turn-Taking, Interruptions, and Task Resolution Without Mistaking Smooth Dialogue for a Reliable Agent

A guide to evaluating GPT‑Live‑1 in full-duplex voice conversations by separating turn dynamics, comprehension, task success, and the safety of delegated actions. It proposes a reproducible protocol using controlled audio, traces, temporal annotation, and blinded human review.

23 Sep 2026
Voxtral TTS: what a team should validate before replacing a synthetic voice in production
NOTICIA

Voxtral TTS: what a team should validate before replacing a synthetic voice in production

Mistral AI documents Voxtral TTS as a speech-generation service with streaming and several output formats. Before replacing a provider or model in calls, notifications, or published content, a team should verify not only audible quality, but also technical, operational, and data-contract compatibility through a regression suite using its own scripts.

22 Sep 2026
GPT-Transcribe: what a team should revalidate before replacing its transcription system
NOTICIA

GPT-Transcribe: what a team should revalidate before replacing its transcription system

GPT-Transcribe is listed as a transcription model available through the OpenAI API. Before replacing an existing ASR system, teams should validate not only the resulting text, but also the technical contract that supports search, summaries, alerts, quotations, reviews, and compliance records.

22 Sep 2026
Synthetic Voices in Production: How to Demonstrate Consent, Control Likeness, and Remove a Voice Without Losing Traceability
GUIA

Synthetic Voices in Production: How to Demonstrate Consent, Control Likeness, and Remove a Voice Without Losing Traceability

Using a synthetic voice responsibly requires more than a checkbox or a convincing demo. This guide proposes an operating system for linking each audio asset to a specific authorization, controlling the risk of resemblance to identifiable people, disclosing its nature where appropriate, and removing assets when the terms of use change.

22 Sep 2026
07 / SOURCES

Traceability