Z
Z.AI / glm-5.1

GLM‑5.1

Coding, research, and professional tool-based workflows.

01 / SUMMARY

The essential

ORIENTATIONCoding and agents

Primary use stated or inferred cautiously from official documentation.

CONTEXT200.000 tokens

Maximum input capacity when published by the source.

EXIT128,000 tokens

Documented maximum generation limit.

ACCESSZ.AI API · open weights

Verified availability channels.

MODEL TYPEText and reasoning

Primary functional family and verified specialties.

LICENSEProprietary

Declared terms for API access, model weights, or self-hosting.

02 / TECHNICAL SHEET

Limits and integration

API IDglm-5.1
Model typeText and reasoning
Access modelPaid proprietary
LicenseProprietary
DeploymentHosted API
Release7 APR 2026
Knowledge cutoffNot published
EntranceText
ExitText
Context window200.000 tokens
maximum output128,000 tokens
ReasoningBuilt-in reasoning
Published toolsFeatures · Web search · Code execution
Structured OutingsSupported
Batch processingNot published
Prompt cacheSupported
Fine-tuningOpen weights
Verified platformsZ.AI API · Hugging Face
BEST SUITED FOR

Coding, research, and professional tool-based workflows.

WORTH MONITORING

Pricing depends on the inference provider or your own hardware.

03 / CAPABILITIES

What it can do

01

Orientation

Coding, research, and professional tool-based workflows.

02

Context

200.000 tokens · 128.000 tokens

03

Tools and integration

Features · Web search · Code execution

04

Access

Z.AI API · open weights

Input modalities
Text
04 / PRICES

Documented cost

ConceptWorthUnit/condition
Standard input1,40 USD / 1 M tokensStandard API rate
Cached input0,26 USD / 1 M tokensReading reused prefixes
Cache write or storageNot publishedThe condition varies by provider
Standard output4,40 USD / 1 M tokensMay include reasoning tokens
Batch inputNot publishedAsynchronous processing
Batch outputNot publishedAsynchronous processing

Consult the primary source before budgeting for a deployment.

i Prices change and may depend on level, region or context length. Check the source before making a decision.

05 / EVALUATIONS

How to read the results

“

A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.

BenchmarkResultMetricSource
SWE-Bench Pro58,4 %ResolvedView source ↗
KernelBench Level 33,6×Geometric mean speedupView source ↗

Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.

06 / VERSIONS

Chronology

GLM‑5.1

Version added to Inferama's verified catalog.

07 / SOURCES

Traceability