Skip to main content
Embedding and Reranking

text-embedding-v4

The General Text Vector V4 version is a multi-language text vector model developed by the Tongyi Lab based on Qwen3. Compared to the V3 version, it significantly improves performance in text retrieval, clustering, and classification tasks. It achieves a 15% to 40% improvement in evaluation tasks such as MTEB multilingual, Chinese-English, and code retrieval. Additionally, it supports user-defined vector dimensions ranging from 64 to 2048.

Inference Service Provider

The inference service provider for text-embedding-v4 is Alibaba Cloud Model Studio.

Model Capabilities

  • China (Beijing)
  • Singapore
  • Hong Kong (China)
CapabilitySupportCapabilitySupport

Input Modality

Text

Output Modality

Text

Model Experience

Unsupported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

ParameterValueParameterValue

Max Input Length

Max Output Length

Context Window

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
  • China (Beijing)
  • Singapore
  • Hong Kong (China)
Billing ItemPrice (USD)Unit

Embedding(Batch File)

0.036

Per 1M tokens

Text Input

0.072

Per 1M tokens

Rate Limits

  • China (Beijing)
  • Singapore
  • Hong Kong (China)
ParameterValue

RPM (Requests Per Minute)

1800

TPM (Tokens Per Minute)

1,200,000

Token Plan
Model Playground
  • Audio generation
  • Music generation
Statistics and Monitoring
Asset Center
Support