Primary use stated or inferred cautiously from official documentation.
Claude Opus 5
Model oriented to code and knowledge work, with controls that allow effort to be adjusted to each task.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Limits and integration
| API ID | claude-opus-5 |
|---|---|
| Release | July 24, 2026 |
| Knowledge cutoff | May 2026 |
| Entrance | Text · Image |
| Exit | Text |
| Context window | 1,000,000 tokens |
| maximum output | 128,000 tokens |
| Reasoning | Adaptive thinking · configurable effort |
| Published tools | Use of tools · Web search · Code execution · Computer use |
| Structured Outings | Compatible |
| Batch processing | Compatible · 50% discount |
| Prompt cache | Compatible · up to 90% savings on reads |
| Fine-tuning | Not published in the source consulted |
| Verified platforms | Claude.ai · Claude Code · Claude API · Amazon Bedrock · Google Cloud · Microsoft Foundry |
Agentic coding, broad refactoring, enterprise knowledge, and visual work.
Responses tend to be lengthy, and active thinking can increase output tokens.
What it can do
Software engineering
Designed for complex planning, editing, and code review.
Knowledge work
Analyze documents and produce structured professional deliverables.
Adjustable effort
It allows you to balance depth, latency and consumption according to the order.
Agentive flows
Maintains multi-step tasks and uses external tools.
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | 5.00 USD / 1 M tokens | Standard API rate |
| Cached input | 0.50 USD / 1 M tokens | Reading reused prefixes |
| Cache write or storage | 6.25 USD / 1 M tokens (5 min) · 10.00 USD (1 h) | The condition varies by provider |
| Standard output | 25.00 USD / 1 M tokens | May include reasoning tokens |
| Batch input | 2.50 USD / 1 M tokens | Asynchronous processing |
| Batch output | 12.50 USD / 1 M tokens | Asynchronous processing |
Thinking is billed as output. Fast mode costs twice the base rate.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
Anthropic highlights improvements in code assessments and knowledge work. Inferama does not make the claim of “best result” without equivalent conditions.
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
Claude Opus 5
Release and general availability.
Claude Opus 4.8
Previous version of the Opus family.