All Products
Search
Document Center

:qwen3-vl-plus

Last Updated:Jul 24, 2026

The Qwen3 series VL models effectively integrates thinking and non-thinking modes, achieving world-leading performance in visual agent capabilities on public benchmark datasets such as OS World. This version features comprehensive upgrades in areas like visual coding, spatial perception, and multimodal reasoning, significantly enhancing visual perception and recognition abilities, and supporting the understanding of ultra-long videos.This model version is functionally equivalent to the snapshot model qwen3-vl-plus-2025-12-19.

Model Capabilities

Capability Support Capability Support

Input Modality

Text Image Video

Output Modality

Text

Model Experience

Supported

Function Calling

Supported

Structured Outputs

Supported

Web Search

Unsupported

Prefix Completion

Supported

Context Caching

Supported

Batch Inference

Supported

Fine-tuning

Unsupported

Context Limits

Parameter Value Parameter Value

Max Input Length

260096

Max Output Length

32768

Context Window

262144

Max Input Length (Thinking Mode)

258048

Max Output Length (Thinking Mode)

32768

Max Chain-of-Thought Length

81920

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Input<=32k

Billing Item Price (USD) Unit

Input

0.144

Per 1M tokens

Output

1.434

Per 1M tokens

Input(Implicit Cache)

0.029

Per 1M tokens

Input(Batch File)

0.072

Per 1M tokens

Output(Batch File)

0.717

Per 1M tokens

Explicit Cache Creation

0.18

Per 1M tokens

Explicit Cache Read

0.014

Per 1M tokens

Input(Batch Chat)

0.144

Per 1M tokens

Output(Batch Chat)

1.434

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.216

Per 1M tokens

Output

2.151

Per 1M tokens

Input(Implicit Cache)

0.044

Per 1M tokens

Input(Batch File)

0.108

Per 1M tokens

Output(Batch File)

1.075

Per 1M tokens

Explicit Cache Creation

0.27

Per 1M tokens

Explicit Cache Read

0.022

Per 1M tokens

Input(Batch Chat)

0.216

Per 1M tokens

Output(Batch Chat)

2.151

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.431

Per 1M tokens

Output

4.301

Per 1M tokens

Input(Implicit Cache)

0.087

Per 1M tokens

Input(Batch File)

0.215

Per 1M tokens

Output(Batch File)

2.15

Per 1M tokens

Explicit Cache Creation

0.539

Per 1M tokens

Explicit Cache Read

0.043

Per 1M tokens

Input(Batch Chat)

0.431

Per 1M tokens

Output(Batch Chat)

4.301

Per 1M tokens

Germany (Frankfurt)

Scope: Global

Input<=32k

Billing Item Price (USD) Unit

Input

0.144

Per 1M tokens

Output

1.434

Per 1M tokens

Input(Implicit Cache)

0.029

Per 1M tokens

Explicit Cache Creation

0.18

Per 1M tokens

Explicit Cache Read

0.014

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.216

Per 1M tokens

Output

2.151

Per 1M tokens

Input(Implicit Cache)

0.044

Per 1M tokens

Explicit Cache Creation

0.27

Per 1M tokens

Explicit Cache Read

0.022

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.431

Per 1M tokens

Output

4.301

Per 1M tokens

Input(Implicit Cache)

0.087

Per 1M tokens

Explicit Cache Creation

0.539

Per 1M tokens

Explicit Cache Read

0.043

Per 1M tokens

US (Virginia)

Scope: Global

Input<=32k

Billing Item Price (USD) Unit

Input

0.144

Per 1M tokens

Output

1.434

Per 1M tokens

Input(Implicit Cache)

0.029

Per 1M tokens

Explicit Cache Creation

0.18

Per 1M tokens

Explicit Cache Read

0.014

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.216

Per 1M tokens

Output

2.151

Per 1M tokens

Input(Implicit Cache)

0.044

Per 1M tokens

Explicit Cache Creation

0.27

Per 1M tokens

Explicit Cache Read

0.022

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.431

Per 1M tokens

Output

4.301

Per 1M tokens

Input(Implicit Cache)

0.087

Per 1M tokens

Explicit Cache Creation

0.539

Per 1M tokens

Explicit Cache Read

0.043

Per 1M tokens

Rate Limits

China (Beijing)

Parameter Value

RPM (Requests Per Minute)

3000

TPM (Tokens Per Minute)

5,000,000

Germany (Frankfurt)

Scope: Global

Parameter Value

RPM (Requests Per Minute)

3000

TPM (Tokens Per Minute)

5,000,000

US (Virginia)

Scope: Global

Parameter Value

RPM (Requests Per Minute)

3000

TPM (Tokens Per Minute)

5,000,000

Snapshot Versions

qwen3-vl-plus-2025-12-19

The Qwen3 series of visual understanding models effectively integrates thinking and non-thinking modes. Compared to the snapshot released on September 23, this version delivers superior performance in reasoning and analysis tasks as well as style control, while also offering lower latency and faster response speeds. This version is based on a snapshot taken on December 19, 2025.

Model Capabilities

Capability Support Capability Support

Input Modality

Text Image Video

Output Modality

Text

Model Experience

Supported

Function Calling

Supported

Structured Outputs

Supported

Web Search

Unsupported

Prefix Completion

Supported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

Parameter Value Parameter Value

Max Input Length

260096

Max Output Length

32768

Context Window

262144

Max Input Length (Thinking Mode)

258048

Max Output Length (Thinking Mode)

32768

Max Chain-of-Thought Length

81920

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Input<=32k

Billing Item Price (USD) Unit

Input

0.144

Per 1M tokens

Output

1.434

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.216

Per 1M tokens

Output

2.151

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.431

Per 1M tokens

Output

4.301

Per 1M tokens

Rate Limits

China (Beijing)

Parameter Value

RPM (Requests Per Minute)

60

TPM (Tokens Per Minute)

100,000

qwen3-vl-plus-2025-09-23

The Qwen3 series VL models effectively integrates thinking and non-thinking modes, achieving world-leading performance in visual agent capabilities on public benchmark datasets such as OS World. This version features comprehensive upgrades in areas like visual coding, spatial perception, and multimodal reasoning, significantly enhancing visual perception and recognition abilities, and supporting the understanding of ultra-long videos.This version is a snapshot as of September 23, 2025

Model Capabilities

Capability Support Capability Support

Input Modality

Text Image Video

Output Modality

Text

Model Experience

Supported

Function Calling

Supported

Structured Outputs

Supported

Web Search

Unsupported

Prefix Completion

Supported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

Parameter Value Parameter Value

Max Input Length

260096

Max Output Length

32768

Context Window

262144

Max Input Length (Thinking Mode)

258048

Max Output Length (Thinking Mode)

32768

Max Chain-of-Thought Length

81920

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Input<=32k

Billing Item Price (USD) Unit

Input

0.143

Per 1M tokens

Output

1.434

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.215

Per 1M tokens

Output

2.15

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.43

Per 1M tokens

Output

4.301

Per 1M tokens

Germany (Frankfurt)

Scope: Global

Input<=32k

Billing Item Price (USD) Unit

Input

0.143

Per 1M tokens

Output

1.434

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.215

Per 1M tokens

Output

2.15

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.43

Per 1M tokens

Output

4.301

Per 1M tokens

US (Virginia)

Scope: Global

Input<=32k

Billing Item Price (USD) Unit

Input

0.143

Per 1M tokens

Output

1.434

Per 1M tokens

32k<Input<=128k

Billing Item Price (USD) Unit

Input

0.215

Per 1M tokens

Output

2.15

Per 1M tokens

128k<Input<=256k

Billing Item Price (USD) Unit

Input

0.43

Per 1M tokens

Output

4.301

Per 1M tokens

Rate Limits

China (Beijing)

Parameter Value

RPM (Requests Per Minute)

60

TPM (Tokens Per Minute)

100,000

Germany (Frankfurt)

Scope: Global

Parameter Value

RPM (Requests Per Minute)

60

TPM (Tokens Per Minute)

100,000

US (Virginia)

Scope: Global

Parameter Value

RPM (Requests Per Minute)

60

TPM (Tokens Per Minute)

100,000