> ## Documentation Index
> Fetch the complete documentation index at: https://docs.modelstudio.console.alibabacloud.com/llms.txt
> Use this file to discover all available pages before exploring further.

# qwen3-asr-flash-realtime

> The real-time version of Qwen3-ASR-Flash is a highly accurate, intelligent, and robust multilingual speech recognition model based on a large language model. Leveraging a powerful foundational model, massive amounts of text and multimodal data, and tens of millions of hours of audio data, Qwen3-ASR-Flash achieves highly accurate speech recognition, automatically determining the language and accurately identifying speech in 11 languages, while ensuring precise transcription even in complex audio environments.This model version is functionally equivalent to the snapshot model qwen3-asr-flash-realtime-2025-10-27.

## Inference Service Provider <span id="h-490b2834e6" />

The inference service provider for `qwen3-asr-flash-realtime` is Alibaba Cloud Model Studio.

## Model Capabilities <span id="h-77230e94b5" />

<table><thead><tr><th>Capability</th><th>Support</th><th>Capability</th><th>Support</th></tr></thead><tbody><tr><td><p>Input Modality</p></td><td><p><strong>Audio</strong></p></td><td><p>Output Modality</p></td><td><p><strong>Text</strong></p></td></tr><tr><td><p>Model Experience</p></td><td><p>Unsupported</p></td><td><p>Function Calling</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Structured Outputs</p></td><td><p>Unsupported</p></td><td><p>Web Search</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Prefix Completion</p></td><td><p>Unsupported</p></td><td><p>Context Caching</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Batch Inference</p></td><td><p>Unsupported</p></td><td><p>Fine-tuning</p></td><td><p>Unsupported</p></td></tr></tbody></table>

## Context Limits <span id="h-01b911f555" />

<table><thead><tr><th>Parameter</th><th>Value</th><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>Max Input Length</p></td><td><p>—</p></td><td><p>Max Output Length</p></td><td><p>—</p></td></tr><tr><td><p>Context Window</p></td><td><p>—</p></td><td><p /></td><td><p /></td></tr></tbody></table>

## Pricing <span id="h-64ff4bce85" />

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit [Model Studio Console](https://modelstudio.console.alibabacloud.com/ap-southeast-1/model/market) for promotional offers.

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Audio Duration</p></td><td><p>0.000047</p></td><td><p>Per second</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Audio Duration</p></td><td><p>0.00009</p></td><td><p>Per second</p></td></tr></tbody></table>
  </Tab>
</Tabs>

## Rate Limits <span id="h-5bab6435a6" />

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>1200</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>1200</p></td></tr></tbody></table>
  </Tab>
</Tabs>

## Snapshot Versions <span id="h-ccf65d0e30" />

### qwen3-asr-flash-realtime-2026-02-10 <span id="h-b14e8b1b13" />

The real-time version of Qwen3-ASR-Flash is a highly accurate, intelligent, and robust multilingual speech recognition model based on a large language model. Leveraging a powerful foundational model, massive amounts of text and multimodal data, and tens of millions of hours of audio data, Qwen3-ASR-Flash achieves highly accurate speech recognition, automatically determining the language and accurately identifying speech in multiple languages, while ensuring precise transcription even in complex audio environments.This version is a snapshot dated February 10, 2026.

#### Inference Service Provider <span id="h-3633db34ea" />

The inference service provider for `qwen3-asr-flash-realtime-2026-02-10` is Alibaba Cloud Model Studio.

#### Model Capabilities <span id="h-1d70300532" />

<table><thead><tr><th>Capability</th><th>Support</th><th>Capability</th><th>Support</th></tr></thead><tbody><tr><td><p>Input Modality</p></td><td><p><strong>Audio</strong></p></td><td><p>Output Modality</p></td><td><p><strong>Text</strong></p></td></tr><tr><td><p>Model Experience</p></td><td><p>Unsupported</p></td><td><p>Function Calling</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Structured Outputs</p></td><td><p>Unsupported</p></td><td><p>Web Search</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Prefix Completion</p></td><td><p>Unsupported</p></td><td><p>Context Caching</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Batch Inference</p></td><td><p>Unsupported</p></td><td><p>Fine-tuning</p></td><td><p>Unsupported</p></td></tr></tbody></table>

#### Context Limits <span id="h-b3b3128fcb" />

<table><thead><tr><th>Parameter</th><th>Value</th><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>Max Input Length</p></td><td><p>—</p></td><td><p>Max Output Length</p></td><td><p>—</p></td></tr><tr><td><p>Context Window</p></td><td><p>—</p></td><td><p /></td><td><p /></td></tr></tbody></table>

#### Pricing <span id="h-ce3b67213e" />

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit [Model Studio Console](https://modelstudio.console.alibabacloud.com/ap-southeast-1/model/market) for promotional offers.

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Audio Duration</p></td><td><p>0.000047</p></td><td><p>Per second</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Audio Duration</p></td><td><p>0.00009</p></td><td><p>Per second</p></td></tr></tbody></table>
  </Tab>
</Tabs>

#### Rate Limits <span id="h-e3a81300c9" />

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>1200</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>1200</p></td></tr></tbody></table>
  </Tab>
</Tabs>

### qwen3-asr-flash-realtime-2025-10-27 <span id="h-3800fa0a33" />

The real-time version of Qwen3-ASR-Flash is a highly accurate, intelligent, and robust multilingual speech recognition model based on a large language model. Leveraging a powerful foundational model, massive amounts of text and multimodal data, and tens of millions of hours of audio data, Qwen3-ASR-Flash achieves highly accurate speech recognition, automatically determining the language and accurately identifying speech in Multiple languages, while ensuring precise transcription even in complex audio environments.This version is a snapshot version from October 27, 2025.

#### Inference Service Provider <span id="h-8a8fa581f9" />

The inference service provider for `qwen3-asr-flash-realtime-2025-10-27` is Alibaba Cloud Model Studio.

#### Model Capabilities <span id="h-b417293654" />

<table><thead><tr><th>Capability</th><th>Support</th><th>Capability</th><th>Support</th></tr></thead><tbody><tr><td><p>Input Modality</p></td><td><p><strong>Audio</strong></p></td><td><p>Output Modality</p></td><td><p><strong>Text</strong></p></td></tr><tr><td><p>Model Experience</p></td><td><p>Unsupported</p></td><td><p>Function Calling</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Structured Outputs</p></td><td><p>Unsupported</p></td><td><p>Web Search</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Prefix Completion</p></td><td><p>Unsupported</p></td><td><p>Context Caching</p></td><td><p>Unsupported</p></td></tr><tr><td><p>Batch Inference</p></td><td><p>Unsupported</p></td><td><p>Fine-tuning</p></td><td><p>Unsupported</p></td></tr></tbody></table>

#### Context Limits <span id="h-a33829c7c0" />

<table><thead><tr><th>Parameter</th><th>Value</th><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>Max Input Length</p></td><td><p>—</p></td><td><p>Max Output Length</p></td><td><p>—</p></td></tr><tr><td><p>Context Window</p></td><td><p>—</p></td><td><p /></td><td><p /></td></tr></tbody></table>

#### Pricing <span id="h-b3fba1d0b6" />

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit [Model Studio Console](https://modelstudio.console.alibabacloud.com/ap-southeast-1/model/market) for promotional offers.

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Audio Duration</p></td><td><p>0.000047</p></td><td><p>Per second</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Billing Item</th><th>Price (USD)</th><th>Unit</th></tr></thead><tbody><tr><td><p>Audio Duration</p></td><td><p>0.00009</p></td><td><p>Per second</p></td></tr></tbody></table>
  </Tab>
</Tabs>

#### Rate Limits <span id="h-e9f8ac7603" />

<Tabs>
  <Tab title="China (Beijing)">
    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>1200</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table><thead><tr><th>Parameter</th><th>Value</th></tr></thead><tbody><tr><td><p>RPM (Requests Per Minute)</p></td><td><p>1200</p></td></tr></tbody></table>
  </Tab>
</Tabs>
