Introduces 'All-in-One Reference' supporting 3-8s videos or multi-image anchoring for character elements. Synchronizes voice dubbing and lip movements for authentic character portrayal. Enhances video consistency and dynamic expression with audio-visual synchronization and intelligent scene segmentation.
Inference Service Provider
The inference service provider for kling/kling-v3-omni-video-generation is Kling AI.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Image Text Video | Output Modality | Video |
Model Experience | Unsupported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | — | Max Output Length | — |
Context Window | — |