Qwen-VL_OCR is an OCR model trained based on Qwen-VL. It aggregates various image-text recognition, parsing, and processing tasks through a unified model approach, offering powerful image-text recognition capabilities.This model version is functionally equivalent to the snapshot model qwen-vl-ocr-2025-11-20.
Inference Service Provider
The inference service provider for qwen-vl-ocr is Alibaba Cloud Model Studio.
Model Capabilities
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Text Image | Output Modality | Text |
Model Experience | Supported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Supported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 30000 | Max Output Length | 8192 |
Context Window | 38192 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.043 | Per 1M tokens |
Output | 0.072 | Per 1M tokens |
Input(Batch File) | 0.022 | Per 1M tokens |
Output(Batch File) | 0.036 | Per 1M tokens |
Rate Limits
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 600 |
TPM (Tokens Per Minute) | 6,000,000 |
Dynamic Updates
qwen-vl-ocr-latest
VL-OCR(qwen-vl-ocr), Qwen-VLOCR., , , .
Inference Service Provider
The inference service provider for qwen-vl-ocr-latest is Alibaba Cloud Model Studio.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Text Image | Output Modality | Text |
Model Experience | Supported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Supported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 30000 | Max Output Length | 8192 |
Context Window | 38192 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.043 | Per 1M tokens |
Output | 0.072 | Per 1M tokens |
Rate Limits
- China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 6000 |
TPM (Tokens Per Minute) | 30,000,000 |
Snapshot Versions
qwen-vl-ocr-2025-11-20
This model is a snapshot version from November 20, 2025, and is based on the latest Qwen-VL3 architecture with a comprehensive upgrade. It features significant improvements in document parsing and text localization capabilities, as well as substantial reductions in end-to-end latency and illusions.
Inference Service Provider
The inference service provider for qwen-vl-ocr-2025-11-20 is Alibaba Cloud Model Studio.
Model Capabilities
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Text Image | Output Modality | Text |
Model Experience | Supported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 30000 | Max Output Length | 8192 |
Context Window | 38192 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.043 | Per 1M tokens |
Output | 0.072 | Per 1M tokens |
Rate Limits
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 6000 |
TPM (Tokens Per Minute) | 30,000,000 |
qwen-vl-ocr-2025-04-13
VL-OCR(qwen-vl-ocr-2025-04-13), Qwen-VLOCR., , , .:, , , , , .20250413.
Inference Service Provider
The inference service provider for qwen-vl-ocr-2025-04-13 is Alibaba Cloud Model Studio.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Text Image | Output Modality | Text |
Model Experience | Supported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 30000 | Max Output Length | 4096 |
Context Window | 34096 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.717 | Per 1M tokens |
Output | 0.717 | Per 1M tokens |
Rate Limits
- China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 600 |
TPM (Tokens Per Minute) | 6,000,000 |
qwen-vl-ocr-1028
VL-OCR(qwen-vl-ocr), Qwen-VLOCR., , , .20241028.
Inference Service Provider
The inference service provider for qwen-vl-ocr-1028 is Alibaba Cloud Model Studio.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Text Image | Output Modality | Text |
Model Experience | Supported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 30000 | Max Output Length | 4096 |
Context Window | 34096 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.717 | Per 1M tokens |
Output | 0.717 | Per 1M tokens |
Rate Limits
- China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 600 |
TPM (Tokens Per Minute) | 6,000,000 |