Character Video Generation model that synthesizes lip-synced videos based on input character videos and voice audio, matching mouth movements to the audio content.
Inference Service Provider
The inference service provider for videoretalk is Alibaba Cloud Model Studio.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Video Audio | Output Modality | Video |
Model Experience | Unsupported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | — | Max Output Length | — |
Context Window | — |
Rate Limits
- China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 60 |