Primary use stated or inferred cautiously from official documentation.
Qwen3‑Max Thinking
Multimodal reasoning, coding, and complex problems.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Primary functional family and verified specialties.
Declared terms for API access, model weights, or self-hosting.
Limits and integration
| API ID | qwen3-max-2026-01-23 |
|---|---|
| Model type | Text and reasoning |
| Access model | Paid proprietary |
| License | Proprietary |
| Deployment | Hosted API |
| Release | 25 JAN 2026 |
| Knowledge cutoff | Not published |
| Entrance | Text |
| Exit | Text |
| Context window | 262.000 tokens |
| maximum output | 65.000 tokens |
| Reasoning | Modes with and without reasoning |
| Published tools | Features · Web search · Web extraction · Code interpreter |
| Structured Outings | Supported |
| Batch processing | Supported |
| Prompt cache | Supported |
| Fine-tuning | Not published |
| Verified platforms | Qwen API · Qwen Chat · Alibaba Cloud |
Multimodal reasoning, coding, and complex problems.
Execution conditions vary across benchmarks; do not compare the number alone.
What it can do
Orientation
Multimodal reasoning, coding, and complex problems.
Context
262.000 tokens · 65.000 tokens
Tools and integration
Features · Web search · Web extraction · Code interpreter
Access
Qwen API · Alibaba Cloud Model Studio
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | 1,20 USD / 1 M tokens (≤32k) | Standard API rate |
| Cached input | 0,24 USD / 1 M tokens | Reading reused prefixes |
| Cache write or storage | Not published | The condition varies by provider |
| Standard output | 6,00 USD / 1 M tokens (≤32k) | May include reasoning tokens |
| Batch input | 0,60 USD / 1 M tokens (≤32k) | Asynchronous processing |
| Batch output | 3,00 USD / 1 M tokens (≤32k) | Asynchronous processing |
Higher pricing tiers apply above 32,000 tokens.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.
| Benchmark | Result | Metric | Source |
|---|---|---|---|
| GPQA Diamond | 92,8 % | Accuracy | View source ↗ |
| LiveCodeBench v6 | 91,4 % | Accuracy | View source ↗ |
| Humanity's Last Exam | 58,3 % | With tools | View source ↗ |
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
Qwen3‑Max Thinking
Version added to Inferama's verified catalog.