O
OPENAI / gpt-oss-120b

gpt‑oss‑120b

Reasoning model with open weights for local or private execution; fits on one H100 GPU according to OpenAI.

01 / SUMMARY

The essential

ORIENTATIONOpen reasoning

Primary use stated or inferred cautiously from official documentation.

CONTEXT131,072 tokens

Maximum input capacity when published by the source.

EXITNot published

Documented maximum generation limit.

ACCESSWeight download

Verified availability channels.

MODEL TYPEText and reasoning · Code

Primary functional family and verified specialties.

LICENSEApache 2.0

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDgpt-oss-120b
Model typeText and reasoning · Code
Access modelOpen weights
LicenseApache 2.0
DeploymentLocal, private cloud, or provider
ReleaseAGO 2025
Knowledge cutoffNot published
EntranceText
ExitText
Context window131,072 tokens
maximum outputNot published
ReasoningAdjustable reasoning
Published tools
Structured OutingsNot applicable or not published
Batch processingNot published
Prompt cacheNot published
Fine-tuningDownloadable weights
Verified platformsHugging Face · Proveedores de inferencia
BEST SUITED FOR

Private reasoning, customization, and deployments that require infrastructure control.

WORTH MONITORING

Requires your own hardware and operations; performance depends on the runtime and quantization.

03 / CAPABILITIES

What it can do

01

Orientation

Private reasoning, customization, and deployments that require infrastructure control.

02

Context

131,072 tokens · Not published

03

Tools and integration

Not published

04

Access

Weight download

Input modalities
Text
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard inputNot publishedStandard API rate
Cached inputNot publishedReading reused prefixes
Cache write or storageNot publishedThe condition varies by provider
Standard outputNot publishedMay include reasoning tokens
Batch inputNot publishedAsynchronous processing
Batch outputNot publishedAsynchronous processing

Consult the primary source: the billing unit depends on the model type and access channel.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

BenchmarkResultMetricSource
Evaluaciones gpt‑ossPublished o4-mini levelGeneral reasoningView source ↗

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

gpt‑oss‑120b

Version added to Inferama's verified catalog.

07 / SOURCES

Traceability