O
OPENAI / gpt-4.1

GPT‑4.1

Long-context non-reasoning model, useful as a historical reference for code, instruction following, and predictable latency.

01 / SUMMARY

The essential

ORIENTATIONNon-reasoning · long context

Primary use stated or inferred cautiously from official documentation.

CONTEXT1,047,576 tokens

Maximum input capacity when published by the source.

EXIT32,768 tokens

Documented maximum generation limit.

ACCESSOpenAI API

Verified availability channels.

MODEL TYPECode · Text and reasoning

Primary functional family and verified specialties.

LICENSEProprietary

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDgpt-4.1
Model typeCode · Text and reasoning
Access modelPaid proprietary
LicenseProprietary
DeploymentHosted API
Release14 ABR 2025
Knowledge cutoffNot published
EntranceText · Image
ExitText
Context window1,047,576 tokens
maximum output32,768 tokens
ReasoningWithout configurable reasoning
Published toolsFeatures · Responses API tools
Structured OutingsSupported
Batch processingNot published
Prompt cacheNot published
Fine-tuningNot published
Verified platformsOpenAI API
BEST SUITED FOR

Code and instruction following when a non-reasoning model is needed.

WORTH MONITORING

OpenAI positions it behind GPT‑5 generations; review its currency before a new integration.

03 / CAPABILITIES

What it can do

01

Orientation

Code and instruction following when a non-reasoning model is needed.

02

Context

1,047,576 tokens · 32,768 tokens

03

Tools and integration

Functions · Responses API tools

04

Access

OpenAI API

Input modalities
TextImage
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard inputNot publishedStandard API rate
Cached inputNot publishedReading reused prefixes
Cache write or storageNot publishedThe condition varies by provider
Standard outputNot publishedMay include reasoning tokens
Batch inputNot publishedAsynchronous processing
Batch outputNot publishedAsynchronous processing

Consult the primary source: the billing unit depends on the model type and access channel.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

BenchmarkResultMetricSource
SWE-bench Verified54,6 %Published solved problemsView source ↗

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

GPT‑4.1

Version added to Inferama's verified catalog.

07 / SOURCES

Traceability