kimi-k2.7-code is Kimi's most intelligent coding model to date. It follows instructions more reliably over long contexts and completes programming tasks with higher success rates. It supports text, image, and video inputs, along with thinking mode, conversation, and agent tasks.
Inference Service Provider
The inference service provider for kimi-k2.7-code is Alibaba Cloud Model Studio.
Model Capabilities
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Text Image Video | Output Modality | Text |
Model Experience | Supported | Function Calling | Supported |
Structured Outputs | Supported | Web Search | Supported |
Prefix Completion | Supported | Context Caching | Supported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | 229376 | Max Output Length | 16384 |
Context Window | 262144 |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Input | 0.8939 | Per 1M tokens |
Output | 3.7131 | Per 1M tokens |
Input(Implicit Cache) | 0.1788 | Per 1M tokens |
Explicit Cache Creation | 1.1174 | Per 1M tokens |
Explicit Cache Read | 0.0894 | Per 1M tokens |
Rate Limits
- China (Beijing)
- Singapore
- Germany (Frankfurt)
- US (Virginia)
- Japan (Tokyo)
- Hong Kong (China)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 500 |
TPM (Tokens Per Minute) | 1,000,000 |