Primary use stated or inferred cautiously from official documentation.
GPT‑6 Astra
Generalist frontier model oriented towards complex tasks, use of tools and highly demanding professional work.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Limits and integration
| API ID | gpt-6-astra |
|---|---|
| Release | September 3, 2026 |
| Knowledge cutoff | April 30, 2026 |
| Entrance | Text · Image |
| Exit | Text |
| Context window | 1,050,000 tokens |
| maximum output | 128,000 tokens |
| Reasoning | low, medium, high, xhigh, and max effort |
| Published tools | Features · Web search · File search · Code interpreter · Hosted shell · Computer use · MCP |
| Structured Outings | Compatible |
| Batch processing | Compatible · 50% discount |
| Prompt cache | Compatible |
| Fine-tuning | Not compatible |
| Verified platforms | OpenAI API |
Complex reasoning, research, code, documents, and agents with tools.
Higher price per token; requires minimum permissions and review for sensitive actions.
What it can do
Advanced reasoning
Breaks down complex objectives and maintains lengthy work plans.
Use of tools
Operate in flows with navigation, code and professional environments.
Multimodal work
Interpret text and visual information within the same task.
Cybersecurity
High capacity accompanied by reinforced safeguards.
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | 10.00 USD / 1 M tokens | Standard API rate |
| Cached input | 1.00 USD / 1 M tokens | Reading reused prefixes |
| Cache write or storage | 12.50 USD / 1 M tokens | The condition varies by provider |
| Standard output | 50.00 USD / 1 M tokens | May include reasoning tokens |
| Batch input | 5.00 USD / 1 M tokens | Asynchronous processing |
| Batch output | 25.00 USD / 1 M tokens | Asynchronous processing |
Requests with more than 272,000 input tokens apply price multipliers; tools may add per-call charges.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
OpenAI presents it as its highest-capability model for complex end-to-end work. The published results are from the provider and should be validated using your own tasks.
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
GPT‑6 Astra
Launch and publication of the security summary.