All Products
Search
Document Center

Alibaba Cloud Model Studio:text-embedding-v4

Last Updated:Sep 28, 2026

The General Text Vector V4 version is a multi-language text vector model developed by the Tongyi Lab based on Qwen3. Compared to the V3 version, it significantly improves performance in text retrieval, clustering, and classification tasks. It achieves a 15% to 40% improvement in evaluation tasks such as MTEB multilingual, Chinese-English, and code retrieval. Additionally, it supports user-defined vector dimensions ranging from 64 to 2048.

Inference Service Provider

The inference service provider for text-embedding-v4 is Alibaba Cloud Model Studio.

Model Capabilities

China (Beijing)

CapabilitySupportCapabilitySupport

Input Modality

Text

Output Modality

Text

Model Experience

Unsupported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Singapore

Scope: International

CapabilitySupportCapabilitySupport

Input Modality

Text

Output Modality

Text

Model Experience

Unsupported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Hong Kong (China)

Scope: Hong Kong

CapabilitySupportCapabilitySupport

Input Modality

Text

Output Modality

Text

Model Experience

Unsupported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

ParameterValueParameterValue

Max Input Length

—

Max Output Length

—

Context Window

—

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.

China (Beijing)

Billing ItemPrice (USD)Unit

Embedding(Batch File)

0.036

Per 1M tokens

Text Input

0.072

Per 1M tokens

Singapore

Scope: International

Billing ItemPrice (USD)Unit

Text Input

0.07

Per 1M tokens

Hong Kong (China)

Scope: Hong Kong

Billing ItemPrice (USD)Unit

Text Input

0.07

Per 1M tokens

Rate Limits

China (Beijing)

ParameterValue

RPM (Requests Per Minute)

1800

TPM (Tokens Per Minute)

1,200,000

Singapore

Scope: International

ParameterValue

RPM (Requests Per Minute)

1800

TPM (Tokens Per Minute)

1,000,000

Hong Kong (China)

Scope: Hong Kong

ParameterValue

RPM (Requests Per Minute)

1800

TPM (Tokens Per Minute)

1,200,000