A lightweight real-time Mandarin speech recognition model optimized for Chinese call-center scenarios, supporting multi-dialect accents and achieving low-latency, high-accuracy transcription in low-sample-rate/low-SNR environments.This model version is functionally equivalent to the snapshot model fun-asr-flash-8k-realtime-2026-01-28.
Inference Service Provider
The inference service provider for fun-asr-flash-8k-realtime is Alibaba Cloud Model Studio.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Audio | Output Modality | Text |
Model Experience | Unsupported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | — | Max Output Length | — |
Context Window | — |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Audio Duration | 0.000032 | Per second |
Rate Limits
- China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 1200 |
Snapshot Versions
fun-asr-flash-8k-realtime-2026-01-28
A lightweight real-time Mandarin speech recognition model optimized for Chinese call-center scenarios, supporting multi-dialect accents and achieving low-latency, high-accuracy transcription in low-sample-rate/low-SNR environments.
Inference Service Provider
The inference service provider for fun-asr-flash-8k-realtime-2026-01-28 is Alibaba Cloud Model Studio.
Model Capabilities
| Capability | Support | Capability | Support |
|---|---|---|---|
Input Modality | Audio | Output Modality | Text |
Model Experience | Unsupported | Function Calling | Unsupported |
Structured Outputs | Unsupported | Web Search | Unsupported |
Prefix Completion | Unsupported | Context Caching | Unsupported |
Batch Inference | Unsupported | Fine-tuning | Unsupported |
Context Limits
| Parameter | Value | Parameter | Value |
|---|---|---|---|
Max Input Length | — | Max Output Length | — |
Context Window | — |
Pricing
This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
- China (Beijing)
| Billing Item | Price (USD) | Unit |
|---|---|---|
Audio Duration | 0.000032 | Per second |
Rate Limits
- China (Beijing)
| Parameter | Value |
|---|---|
RPM (Requests Per Minute) | 1200 |