A 2.4-trillion-parameter MoE flagship with a major leap in coding and office productivity, able to autonomously code for over ten days to deliver complete projects. It handles hundreds of professional tasks across law, finance, design, and more, delivering production-grade results end-to-end in a single conversation. Native visual understanding runs through the entire planning, execution, and verification pipeline, enabling deep semantic parsing of ultra-long documents and long videos. It plans autonomously and iterates in closed loops during long-horizon tasks, continuously improving.
Model Capabilities
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Image Text Video | Output Modality | Text |
Model Experience | Supported | Function Calling | Supported |
Structured Outputs | Supported | Web Search | Supported |
Partial Mode | Supported | Context Cache | Supported |
Batch Inference | Supported | Fine-tuning | Unsupported |
Context Limits
Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 991808 | Max Output Length | 131072 |
Max Input Length (Thinking Mode) | 983616 | Max Output Length (Thinking Mode) | 131072 |
Context Window | 1000000 | Max Chain-of-Thought Length | 262144 |
The supported model length may vary depending on different combinations of API input parameters.
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 1.65 | Per 1M tokens |
Output | 4.951 | Per 1M tokens |
Input(Implicit Cache) | 0.206 | Per 1M tokens |
Explicit Cache Creation | 2.063 | Per 1M tokens |
Explicit Cache Read | 0.137 | Per 1M tokens |
Input(Batch File) | 0.825 | Per 1M tokens |
Output(Batch File) | 2.475 | Per 1M tokens |
Input(Batch Chat) | 1.65 | Per 1M tokens |
Output(Batch Chat) | 4.951 | Per 1M tokens |
Rate Limits
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 30,000 |
TPM (Tokens Per Minute) | 5,000,000 |