> ## Documentation Index
> Fetch the complete documentation index at: https://docs.modelstudio.console.alibabacloud.com/llms.txt
> Use this file to discover all available pages before exploring further.

# qwen3-livetranslate-flash

> Qwen3-LiveTranslate-Flash is a high-precision, highly responsive, and robust multilingual real-time audio and video interpretation model. Leveraging Qwen3-Omni’s powerful infrastructure, massive multimodal data, cross-language and cross-modal alignment, and visual enhancement technologies, Qwen3-LiveTranslate-Flash provides both offline and real-time audio and video translation capabilities. It can understand 19 languages and speak 10 languages, and also supports 8 Chinese dialects.

## Inference Service Provider <span id="h-490b2834e6" />

The inference service provider for `qwen3-livetranslate-flash` is Alibaba Cloud Model Studio.

## Model Capabilities <span id="h-77230e94b5" />

<table><thead><tr><th>Capability</th><th>Support</th><th>Capability</th><th>Support</th></tr></thead><tbody><tr><td><p>Input Modality</p></td><td><p><strong>Audio</strong> <strong>Video</strong></p></td><td><p>Output Modality</p></td><td><p><strong>Text</strong> <strong>Audio</strong></p></td></tr><tr><td><p>Model Experience</p></td><td><p>Unsupported</p></td><td><p>Function Calling</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Structured Outputs</p></td><td><p>Unsupported</p></td><td><p>Web Search</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Prefix Completion</p></td><td><p>Unsupported</p></td><td><p>Context Caching</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Batch Inference</p></td><td><p>Unsupported</p></td><td><p>Fine-tuning</p></td><td><p>Unsupported</p></td></tr></tbody></table>

## Context Limits <span id="h-01b911f555" />

<table><thead><tr><th>Parameter</th><th>Value</th><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>Max Input Length</p></td><td><p>49152</p></td><td><p>Max Output Length</p></td><td><p>4096</p></td></tr><tr><td><p>Context Window</p></td><td><p>53248</p></td><td><p /></td><td><p /></td></tr></tbody></table>

## Pricing <span id="h-fbea8571b2" />

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit [Model Studio Console](https://modelstudio.console.alibabacloud.com/ap-southeast-1/model/market) for promotional offers.

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Input: Audio</p></td><td><p>1.434</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Input: Image</p></td><td><p>0.573</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Output: Text</p></td><td><p>1.434</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Output: Audio</p></td><td><p>5.734</p></td><td><p>Per 1M tokens</p></td></tr></tbody></table>
  </Tab>
</Tabs>

## Rate Limits <span id="h-c38925b202" />

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>100</p></td></tr><tr><td><p>TPM (Tokens Per Minute)</p></td><td><p>100,000</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>100</p></td></tr><tr><td><p>TPM (Tokens Per Minute)</p></td><td><p>100,000</p></td></tr></tbody></table>
  </Tab>
</Tabs>

## Snapshot Versions <span id="h-a3d6551982" />

### qwen3-livetranslate-flash-2025-12-01 <span id="h-861a902ef6" />

Qwen3-LiveTranslate-Flash is a high-precision, highly responsive, and robust multilingual real-time audio and video interpretation model. Leveraging Qwen3-Omni’s powerful infrastructure, massive multimodal data, cross-language and cross-modal alignment, and visual enhancement technologies, Qwen3-LiveTranslate-Flash provides both offline and real-time audio and video translation capabilities. It can understand 19 languages and speak 10 languages, and also supports 8 Chinese dialects. This version is a snapshot from December 1, 2025.

#### Inference Service Provider <span id="h-672edebe26" />

The inference service provider for `qwen3-livetranslate-flash-2025-12-01` is Alibaba Cloud Model Studio.

#### Model Capabilities <span id="h-10dc9f554d" />

<table><thead><tr><th>Capability</th><th>Support</th><th>Capability</th><th>Support</th></tr></thead><tbody><tr><td><p>Input Modality</p></td><td><p><strong>Audio</strong> <strong>Video</strong></p></td><td><p>Output Modality</p></td><td><p><strong>Text</strong> <strong>Audio</strong></p></td></tr><tr><td><p>Model Experience</p></td><td><p>Unsupported</p></td><td><p>Function Calling</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Structured Outputs</p></td><td><p>Unsupported</p></td><td><p>Web Search</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Prefix Completion</p></td><td><p>Unsupported</p></td><td><p>Context Caching</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Batch Inference</p></td><td><p>Unsupported</p></td><td><p>Fine-tuning</p></td><td><p>Unsupported</p></td></tr></tbody></table>

#### Context Limits <span id="h-f493e4fcd6" />

<table><thead><tr><th>Parameter</th><th>Value</th><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>Max Input Length</p></td><td><p>49152</p></td><td><p>Max Output Length</p></td><td><p>4096</p></td></tr><tr><td><p>Context Window</p></td><td><p>53248</p></td><td><p /></td><td><p /></td></tr></tbody></table>

#### Pricing <span id="h-f5382eb139" />

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit [Model Studio Console](https://modelstudio.console.alibabacloud.com/ap-southeast-1/model/market) for promotional offers.

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Input: Audio</p></td><td><p>1.434</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Input: Image</p></td><td><p>0.5734</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Output: Text</p></td><td><p>1.434</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Output: Audio</p></td><td><p>5.734</p></td><td><p>Per 1M tokens</p></td></tr></tbody></table>
  </Tab>
</Tabs>

#### Rate Limits <span id="h-95c1936bfc" />

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>100</p></td></tr><tr><td><p>TPM (Tokens Per Minute)</p></td><td><p>100,000</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>100</p></td></tr><tr><td><p>TPM (Tokens Per Minute)</p></td><td><p>100,000</p></td></tr></tbody></table>
  </Tab>
</Tabs>
