Primary use stated or inferred cautiously from official documentation.
GPT‑Live 1
Real-time voice model for expressive conversations, natural interruptions, and tool-enabled workflows.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Primary functional family and verified specialties.
Declared terms for API access, model weights, or self-hosting.
Limits and integration
| API ID | gpt-live-1 |
|---|---|
| Model type | Voice and audio · Text and reasoning |
| Access model | Paid proprietary |
| License | Proprietary |
| Deployment | Hosted API |
| Release | 2026 |
| Knowledge cutoff | Not published |
| Entrance | Text · Audio |
| Exit | Text · Audio |
| Context window | Real-time session |
| maximum output | Streaming audio and text |
| Reasoning | Not applicable or not published |
| Published tools | Features · Voice sessions |
| Structured Outings | Not applicable or not published |
| Batch processing | Not published |
| Prompt cache | Not published |
| Fine-tuning | Not published |
| Verified platforms | OpenAI Realtime API |
Voice agents, conversational support, and interactive spoken experiences.
Measure end-to-end latency, interruptions, cost per audio, and multilingual behavior.
What it can do
Orientation
Voice agents, conversational support, and interactive spoken experiences.
Context
Real-time session · Streaming audio and text
Tools and integration
Functions · Voice sessions
Access
OpenAI API
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | Not published | Standard API rate |
| Cached input | Not published | Reading reused prefixes |
| Cache write or storage | Not published | The condition varies by provider |
| Standard output | Not published | May include reasoning tokens |
| Batch input | Not published | Asynchronous processing |
| Batch output | Not published | Asynchronous processing |
Consult the primary source: the billing unit depends on the model type and access channel.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.
| Benchmark | Result | Metric | Source |
|---|---|---|---|
| Evaluación de conversación GPT‑Live | Main model | Naturalness and interruptions | View source ↗ |
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
GPT‑Live 1
Version added to Inferama's verified catalog.
Related analysis
GPT‑Live Conversation Evaluation: How to Measure Turn-Taking, Interruptions, and Task Resolution Without Mistaking Smooth Dialogue for a Reliable Agent
A guide to evaluating GPT‑Live‑1 in full-duplex voice conversations by separating turn dynamics, comprehension, task success, and the safety of delegated actions. It proposes a reproducible protocol using controlled audio, traces, temporal annotation, and blinded human review.
23 Sep 2026 ↗ NOTICIAVoxtral TTS: what a team should validate before replacing a synthetic voice in production
Mistral AI documents Voxtral TTS as a speech-generation service with streaming and several output formats. Before replacing a provider or model in calls, notifications, or published content, a team should verify not only audible quality, but also technical, operational, and data-contract compatibility through a regression suite using its own scripts.
22 Sep 2026 ↗ NOTICIAGPT-Transcribe: what a team should revalidate before replacing its transcription system
GPT-Transcribe is listed as a transcription model available through the OpenAI API. Before replacing an existing ASR system, teams should validate not only the resulting text, but also the technical contract that supports search, summaries, alerts, quotations, reviews, and compliance records.
22 Sep 2026 ↗ GUIASynthetic Voices in Production: How to Demonstrate Consent, Control Likeness, and Remove a Voice Without Losing Traceability
Using a synthetic voice responsibly requires more than a checkbox or a convincing demo. This guide proposes an operating system for linking each audio asset to a specific authorization, controlling the risk of resemblance to identifiable people, disclosing its nature where appropriate, and removing assets when the terms of use change.
22 Sep 2026 ↗