O
OPENAI / gpt-5.6-luna

GPT‑5.6 Luna

High volume, subagents, and structured extraction.

01 / SUMMARY

The essential

ORIENTATIONHigh efficiency

Primary use stated or inferred cautiously from official documentation.

CONTEXT1,050,000 tokens

Maximum input capacity when published by the source.

EXIT128,000 tokens

Documented maximum generation limit.

ACCESSOpenAI API · Responses and Chat Completions

Verified availability channels.

MODEL TYPEText and reasoning

Primary functional family and verified specialties.

LICENSEProprietary

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDgpt-5.6-luna
Model typeText and reasoning
Access modelPaid proprietary
LicenseProprietary
DeploymentHosted API
ReleaseJUL 2026
Knowledge cutoff16 FEB 2026
EntranceText · Image
ExitText
Context window1,050,000 tokens
maximum output128,000 tokens
Reasoningnone, low, medium, high, xhigh, and max effort
Published toolsFeatures · Web search · File search · Computer use
Structured OutingsSupported
Batch processingCompatible · 50% discount
Prompt cacheCompatible · 90% discount on reads
Fine-tuningNot published
Verified platformsOpenAI API
BEST SUITED FOR

High volume, subagents, and structured extraction.

WORTH MONITORING

Cost and performance depend on the reasoning level.

03 / CAPABILITIES

What it can do

01

Orientation

High volume, subagents, and structured extraction.

02

Context

1.050.000 tokens · 128.000 tokens

03

Tools and integration

Features · Web search · File search · Computer use

04

Access

OpenAI API · Responses and Chat Completions

Input modalities
TextImage
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard input0,20 USD / 1 M tokensStandard API rate
Cached input0,02 USD / 1 M tokensReading reused prefixes
Cache write or storage0,25 USD / 1 M tokensThe condition varies by provider
Standard output1,20 USD / 1 M tokensMay include reasoning tokens
Batch input0,10 USD / 1 M tokensAsynchronous processing
Batch output0,60 USD / 1 M tokensAsynchronous processing

Consult the primary source before budgeting for a deployment.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

BenchmarkResultMetricSource
Agents' Last Exam50,3 %AccuracyView source ↗
SWE-Bench Pro62,7 %ResolvedView source ↗
Terminal-Bench 2.184,7 %AccuracyView source ↗

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

GPT‑5.6 Luna

Version added to Inferama's verified catalog.

07 / SOURCES

Traceability