The Qwen3.6 native vision-language Plus series models demonstrate exceptional performance on par with the current state-of-the-art models, with a significant improvement in overall results compared to the 3.5 series. The models have been markedly enhanced in code-related capabilities such as agentic coding, front-end programming, and Vibe coding, as well as in multi-modal general object recognition, OCR, and object localization.This model version is functionally equivalent to the snapshot model qwen3.6-plus-2026-04-02.
Inference Service Provider
The inference service provider for qwen3.6-plus is Alibaba Cloud Model Studio.
Model Capabilities
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Image Text Video | Output Modality | Text |
Model Experience | Supported | Function Calling | Supported |
Structured Outputs | Supported | Web Search | Supported |
Prefix Completion | Supported | Context Caching | Supported |
Batch Inference | Supported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 991808 | Max Output Length | 65536 |
Context Window | 1000000 | Max Input Length (Thinking Mode) | 983616 |
Max Output Length (Thinking Mode) | 65536 | Max Chain-of-Thought Length | 81920 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
Input<=256k
Input<=256k
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.276 | Per 1M tokens |
Output | 1.651 | Per 1M tokens |
Input(Batch File) | 0.138 | Per 1M tokens |
Output(Batch File) | 0.825 | Per 1M tokens |
Explicit Cache Creation | 0.344 | Per 1M tokens |
Explicit Cache Read | 0.028 | Per 1M tokens |
Input(Batch Chat) | 0.275 | Per 1M tokens |
Output(Batch Chat) | 1.65 | Per 1M tokens |
256k<Input<=1m
256k<Input<=1m
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 1.101 | Per 1M tokens |
Output | 6.602 | Per 1M tokens |
Input(Batch File) | 0.55 | Per 1M tokens |
Output(Batch File) | 3.301 | Per 1M tokens |
Explicit Cache Creation | 1.376 | Per 1M tokens |
Explicit Cache Read | 0.111 | Per 1M tokens |
Input(Batch Chat) | 1.1 | Per 1M tokens |
Output(Batch Chat) | 6.601 | Per 1M tokens |
Rate Limits
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 30000 |
TPM (Tokens Per Minute) | 5,000,000 |
Snapshot Versions
qwen3.6-plus-2026-04-02
The Qwen3.6 native vision-language Plus series models demonstrate exceptional performance on par with the current state-of-the-art models, with a significant improvement in overall results compared to the 3.5 series. The models have been markedly enhanced in code-related capabilities such as agentic coding, front-end programming, and Vibe coding, as well as in multi-modal general object recognition, OCR, and object localization.This version is a snapshot as of April 2, 2026.
Inference Service Provider
The inference service provider for qwen3.6-plus-2026-04-02 is Alibaba Cloud Model Studio.
Model Capabilities
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Image Text Video | Output Modality | Text |
Model Experience | Supported | Function Calling | Supported |
Structured Outputs | Supported | Web Search | Supported |
Prefix Completion | Supported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 991808 | Max Output Length | 65536 |
Context Window | 1000000 | Max Input Length (Thinking Mode) | 983616 |
Max Output Length (Thinking Mode) | 65536 | Max Chain-of-Thought Length | 81920 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
Input<=256k
Input<=256k
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.276 | Per 1M tokens |
Output | 1.651 | Per 1M tokens |
256k<Input<=1m
256k<Input<=1m
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 1.101 | Per 1M tokens |
Output | 6.602 | Per 1M tokens |
Rate Limits
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 600 |
TPM (Tokens Per Minute) | 1,000,000 |