The Max model, the largest and most capable in the Qwen3.7 series, currently offers a pure‑text‑only interface for public experimentation. Qwen3.7 is a next‑generation flagship model designed for the agent‑centric era, with its core strengths lying in the breadth and depth of its agent‑level capabilities: it excels at programming, office and productivity tasks, and long‑term autonomous execution.This model version is functionally equivalent to the snapshot model qwen3.7-max-2026-05-20.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality |
Text |
Output Modality |
Text |
Model Experience |
Function Calling |
||
Structured Outputs |
Web Search |
||
Prefix Completion |
Context Caching |
||
Batch Inference |
Fine-tuning |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length |
991808 |
Max Output Length |
65536 |
Context Window |
1000000 |
Max Input Length (Thinking Mode) |
983616 |
Max Output Length (Thinking Mode) |
65536 |
Max Chain-of-Thought Length |
262144 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Input(Batch File) |
0.825 |
Per 1M tokens |
Output(Batch File) |
2.475 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Input(Batch Chat) |
1.65 |
Per 1M tokens |
Output(Batch Chat) |
4.951 |
Per 1M tokens |
Germany (Frankfurt)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
US (Virginia)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Japan (Tokyo)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Hong Kong (China)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Rate Limits
China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
30000 |
TPM (Tokens Per Minute) |
5,000,000 |
Germany (Frankfurt)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
30000 |
TPM (Tokens Per Minute) |
5,000,000 |
US (Virginia)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
30000 |
TPM (Tokens Per Minute) |
5,000,000 |
Japan (Tokyo)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
30000 |
TPM (Tokens Per Minute) |
5,000,000 |
Hong Kong (China)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
30000 |
TPM (Tokens Per Minute) |
5,000,000 |
Snapshot Versions
qwen3.7-max-2026-06-08
The Max model, the largest and most capable in the Qwen3.7 series, has added visual‑modal understanding compared to the May 20 snapshot, enabling it to perceive real‑world scenes and supporting multimodal interactive hybrid agent capabilities. This version is based on a snapshot taken on June 8, 2026.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality |
Image Text Video |
Output Modality |
Text |
Model Experience |
Function Calling |
||
Structured Outputs |
Web Search |
||
Prefix Completion |
Context Caching |
||
Batch Inference |
Fine-tuning |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length |
991808 |
Max Output Length |
65536 |
Context Window |
1000000 |
Max Input Length (Thinking Mode) |
983616 |
Max Output Length (Thinking Mode) |
65536 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Germany (Frankfurt)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
US (Virginia)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Hong Kong (China)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Input(Implicit Cache) |
0.33 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Rate Limits
China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
Germany (Frankfurt)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
US (Virginia)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
Hong Kong (China)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
qwen3.7-max-2026-05-20
The Max model, the largest and most capable in the Qwen3.7 series, currently offers a pure‑text‑only interface for public experimentation. Qwen3.7 is a next‑generation flagship model designed for the agent‑centric era, with its core strengths lying in the breadth and depth of its agent‑level capabilities: it excels at programming, office and productivity tasks, and long‑term autonomous execution.This version is a snapshot as of May 20, 2026.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality |
Text |
Output Modality |
Text |
Model Experience |
Function Calling |
||
Structured Outputs |
Web Search |
||
Prefix Completion |
Context Caching |
||
Batch Inference |
Fine-tuning |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length |
991808 |
Max Output Length |
65536 |
Context Window |
1000000 |
Max Input Length (Thinking Mode) |
983616 |
Max Output Length (Thinking Mode) |
65536 |
Max Chain-of-Thought Length |
262144 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Germany (Frankfurt)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
US (Virginia)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Japan (Tokyo)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Hong Kong (China)
Scope: Global
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input |
1.65 |
Per 1M tokens |
Output |
4.951 |
Per 1M tokens |
Explicit Cache Creation |
2.063 |
Per 1M tokens |
Explicit Cache Read |
0.165 |
Per 1M tokens |
Rate Limits
China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
Germany (Frankfurt)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
US (Virginia)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
Japan (Tokyo)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |
Hong Kong (China)
Scope: Global
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) |
600 |
TPM (Tokens Per Minute) |
1,000,000 |