Primary use stated or inferred cautiously from official documentation.
GPT‑6.1 Sol
OpenAI model for complex coding, computer use and professional work with text and image inputs. The provider describes performance close to Astra at a lower cost; assess that trade-off on your own tasks.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Primary functional family and verified specialties.
Declared terms for API access, model weights, or self-hosting.
Limits and integration
| API ID | gpt-6.1-sol |
|---|---|
| Model type | Text and reasoning · Code · Vision and documents |
| Access model | Paid proprietary |
| License | Proprietary |
| Deployment | Hosted API |
| Release | 29 SEP 2026 |
| Knowledge cutoff | 30 ABR 2026 |
| Entrance | Text · Image |
| Exit | Text |
| Context window | 1,050,000 tokens · maximum 922,000 input tokens |
| maximum output | 128,000 tokens |
| Reasoning | low, medium, high, xhigh, and max effort |
| Published tools | Features · Web search · File search · Code execution · Computer use · MCP · Tool search · Multi-agent (beta) |
| Structured Outings | Supported |
| Batch processing | Compatible · 50% discount |
| Prompt cache | Supported · 95% discount on cache reads |
| Fine-tuning | Not compatible |
| Verified platforms | OpenAI API |
Software engineering, agents with tools and long documents.
Tools require the Responses API; Chat Completions does not support tools. The none and minimal reasoning effort levels are not supported. Fast mode is unavailable with EU data residency.
What it can do
Orientation
Software engineering, agents with tools and long documents.
Context
1,050,000 tokens · maximum 922,000 input tokens · 128,000 tokens
Tools and integration
Function calling · Web search · File search · Code execution · Computer use · MCP · Tool search · Multi-agent (beta)
Access
OpenAI API · Responses and Chat Completions
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | 2,00 USD / 1 M tokens | Standard API rate |
| Cached input | 0,10 USD / 1 M tokens | Reading reused prefixes |
| Cache write or storage | 2.50 USD / 1 M tokens | The condition varies by provider |
| Standard output | 10.00 USD / 1 M tokens | May include reasoning tokens |
| Batch input | 1.00 USD / 1 M tokens | Asynchronous processing |
| Batch output | 5.00 USD / 1 M tokens | Asynchronous processing |
Standard prices apply up to 272,000 input tokens. Above that threshold, the entire request costs twice as much for input and cache and 50% more for output. Fast costs twice as much; Batch and Flex reduce prices by 50%. Regional processing adds 10% where available.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.
i We have not found published benchmarks for this model.
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
GPT‑6.1 Sol
Version added to Inferama's verified catalog.