Qwen3 Max
qwen/qwen3-maxAlibaba's closed flagship: broad knowledge, strong multilingual output and vision input at mid-tier pricing.
Model overview
Context window
256K
tokens
Input price
—
quoted on request
Output price
—
quoted on request
Weekly volume
—
tokens / week
Open weights
No
API only
Capabilities
5
of 11 tags
Description
Qwen3 Max is a language model from Alibaba Qwen. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. Alibaba Qwen’s stated focus is the broadest open model family in the world. The weights are not published, so it is API-only.
Capabilities
- Tool calling
- Calls functions you define and returns their arguments as structured data, the basis of agents.
- JSON mode
- Constrains the response to valid JSON matching a schema you supply.
- Streaming
- Emits tokens as they are generated, so answers appear progressively rather than all at once.
- Vision input
- Accepts images alongside text in the prompt.
- Caching
- Reuses already-processed prompt prefixes across requests, cutting input cost for repeated context.
Strengths & weaknesses
Strengths
- Excellent multilingual quality
- Vision input included
- Strong all-round benchmark profile
Weaknesses
- Closed weights unlike the rest of the family
- Output pricing above open peers
Pricing
Free| Rate | Price | Unit |
|---|---|---|
| Input | $0.00 | per 1M tokens |
| Output | — not charged | — |
Prompt caching is supported, so repeated prefixes bill below the listed input rate. Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.
Context window
256Ktokens
covers prompt and response together, so a long input leaves less room for the answer.
Supported features
| Feature | Support |
|---|---|
| Tool calling | Supported |
| JSON mode | Supported |
| Streaming | Supported |
| Vision input | Supported |
| Audio input | Not supported |
| Long context | Not supported |
| Open weights | Not supported |
| Reasoning | Not supported |
| Fine-tunable | Not supported |
| Batch | Not supported |
| Caching | Supported |
Example use cases
Multilingual products
Excellent multilingual quality
Document understanding
Vision input included
General assistants
Strong all-round benchmark profile
Comparison
| Attribute | Qwen3 Max | Qwen3 235B A22B Instruct 2507 | GPT-5.1 |
|---|---|---|---|
| Context window | 256K | 256K | 400K |
| Input price | $0.00 | $0.00 | $0.00 |
| Output price | — not charged | — not charged | — not charged |
| Tool calling | Supported | Supported | Supported |
| JSON mode | Supported | Supported | Supported |
| Vision input | Supported | Not supported | Supported |
| Open weights | Not supported | Supported | Not supported |
| Prompt caching | Supported | Not supported | Supported |
Best value in each row is highlighted. Illustrative placeholder data.
Provider information
View Alibaba Qwen →Alibaba Qwen · 通义千问 · Hangzhou, China · est. 2023 (Model lab; parent Alibaba founded 1999)
Alibaba's Qwen family spans every size class from edge models to frontier MoE systems, most released with open weights. Its breadth (chat, coding, vision, embeddings) and permissive licensing made it the default base model for much of the global open-source ecosystem.
Recent releases
Documentation
Frequently asked questions
Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Prompt caching bills repeated prefixes below the listed input rate. Every figure on this page is an illustrative placeholder.