O
OPENAI / gpt-6.1-sol

GPT‑6.1 Sol

OpenAI model for complex coding, computer use and professional work with text and image inputs. The provider describes performance close to Astra at a lower cost; assess that trade-off on your own tasks.

01 / SUMMARY

The essential

ORIENTATIONCoding, computer use and professional work

Primary use stated or inferred cautiously from official documentation.

CONTEXT1,050,000 tokens · maximum 922,000 input tokens

Maximum input capacity when published by the source.

EXIT128,000 tokens

Documented maximum generation limit.

ACCESSOpenAI API · Responses and Chat Completions

Verified availability channels.

MODEL TYPEText and reasoning · Code · Vision and documents

Primary functional family and verified specialties.

LICENSEProprietary

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDgpt-6.1-sol
Model typeText and reasoning · Code · Vision and documents
Access modelPaid proprietary
LicenseProprietary
DeploymentHosted API
Release29 SEP 2026
Knowledge cutoff30 ABR 2026
EntranceText · Image
ExitText
Context window1,050,000 tokens · maximum 922,000 input tokens
maximum output128,000 tokens
Reasoninglow, medium, high, xhigh, and max effort
Published toolsFeatures · Web search · File search · Code execution · Computer use · MCP · Tool search · Multi-agent (beta)
Structured OutingsSupported
Batch processingCompatible · 50% discount
Prompt cacheSupported · 95% discount on cache reads
Fine-tuningNot compatible
Verified platformsOpenAI API
BEST SUITED FOR

Software engineering, agents with tools and long documents.

WORTH MONITORING

Tools require the Responses API; Chat Completions does not support tools. The none and minimal reasoning effort levels are not supported. Fast mode is unavailable with EU data residency.

03 / CAPABILITIES

What it can do

01

Orientation

Software engineering, agents with tools and long documents.

02

Context

1,050,000 tokens · maximum 922,000 input tokens · 128,000 tokens

03

Tools and integration

Function calling · Web search · File search · Code execution · Computer use · MCP · Tool search · Multi-agent (beta)

04

Access

OpenAI API · Responses and Chat Completions

Input modalities
TextImage
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard input2,00 USD / 1 M tokensStandard API rate
Cached input0,10 USD / 1 M tokensReading reused prefixes
Cache write or storage2.50 USD / 1 M tokensThe condition varies by provider
Standard output10.00 USD / 1 M tokensMay include reasoning tokens
Batch input1.00 USD / 1 M tokensAsynchronous processing
Batch output5.00 USD / 1 M tokensAsynchronous processing

Standard prices apply up to 272,000 input tokens. Above that threshold, the entire request costs twice as much for input and cache and 50% more for output. Fast costs twice as much; Batch and Flex reduce prices by 50%. Regional processing adds 10% where available.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

i We have not found published benchmarks for this model.

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

GPT‑6.1 Sol

Version added to Inferama's verified catalog.

07 / SOURCES

Traceability