Skip to main content
Billing

Model inference pricing

Model API calls are billed on a pay-as-you-go basis by default.
This document only lists standard prices. For the latest promotions, visit the Model Studio console.
Some models support context caching (explicit cache and implicit cache). Cache-hit input tokens and the tokens used to create an explicit cache are billed at unit prices different from the standard input price (for example, explicit cache creation is billed at 125% of the standard input price, and cache hits at 10%). The input prices in the tables below do not include cache prices. For cache billing rules, discount rates, and supported models, see Context Cache.

Tiered pricing rules

Some Model Studio models use tiered pricing. The unit price is determined by the total number of input tokens in a single request. All tokens in the request are billed at the unit price of the corresponding tier. In the pricing tiers, K means 1,000 and M means 1,000,000. For example, 128K equals 128,000 tokens, 256K equals 256,000 tokens, and 1M equals 1,000,000 tokens. For example, a model has two pricing tiers: 0 < tokens ≤ 32K and 32K < tokens ≤ 128K. If a request contains 100K input tokens, it falls into the second tier (32K < 100K ≤ 128K), and all tokens are billed at the unit price of the second tier.

Text generation - Qwen

Qwen-Max

You are charged for input tokens and output tokens. If the model supports batch calls, the unit price for both input and output tokens is 50% of the real-time inference price. If the model supports context cache, only input tokens receive a discount. These two discounts cannot apply simultaneously.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
Model IDDeployment scopeModeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought + answer
Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3.8-max
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤1M$2$61 million tokens
qwen3.7-max
Currently equivalent to qwen3.7-max-2026-05-20
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤1MList price $2.5 (Limited-time 50% off)List price $7.5 (Limited-time 50% off)1 million tokens
qwen3.7-max-2026-06-08
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤1M$2.5$7.51 million tokens
qwen3.7-max-2026-05-20
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤1M$2.5$7.51 million tokens
qwen3.7-max-preview
Currently equivalent to qwen3.7-max-2026-05-17
InternationalThinking mode only0<Token≤1M$2.5$7.51 million tokens
qwen3.7-max-2026-05-17InternationalThinking mode only0<Token≤1M$2.5$7.51 million tokens
qwen3.6-max-preview
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤128K$1.3$7.81 million tokens
128K<Token≤256K$2$12
qwen3-max
Currently equivalent to qwen3-max-2026-01-23
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤32K$1.2$61 million tokens
32K<Token≤128K$2.4$12
128K<Token≤256K$3$15
qwen3-max-2026-01-23InternationalNon-Thinking and Thinking modes0<Token≤32K$1.2$61 million tokens
32K<Token≤128K$2.4$12
128K<Token≤256K$3$15
qwen3-max-2025-09-23InternationalNon-Thinking mode only0<Token≤32K$1.2$61 million tokens
32K<Token≤128K$2.4$12
128K<Token≤256K$3$15
qwen3-max-preview
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤32K$1.2$61 million tokens
32K<Token≤128K$2.4$12
128K<Token≤256K$3$15
More models
Model IDDeployment scopeModeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen-max
50% batch inference discount
InternationalNon-Thinking mode onlyNo tiered pricing$1.6$6.41 million tokens

Qwen-Plus

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
Model IDDeployment scopeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
Non-Thinking modeThinking mode (chain of thought + answer)
qwen3.7-plus
Currently equivalent to qwen3.7-plus-2026-05-26
context caching discount
International0<Token≤256KList price $0.4 (Limited-time 20% off)List price $1.6 (Limited-time 20% off)List price $1.6 (Limited-time 20% off)1 million tokens
256K<Token≤1MList price $1.2 (Limited-time 20% off)List price $4.8 (Limited-time 20% off)List price $4.8 (Limited-time 20% off)
qwen3.7-plus-2026-05-26
context caching discount
International0<Token≤256K$0.4$1.6$1.61 million tokens
256K<Token≤1M$1.2$4.8$4.8
qwen3.6-plus
Currently equivalent to qwen3.6-plus-2026-04-02
International0<Token≤256K$0.5$3$31 million tokens
256K<Token≤1M$2$6$6
qwen3.6-plus-2026-04-02International0<Token≤256K$0.5$3$31 million tokens
256K<Token≤1M$2$6$6
qwen3.5-plus
Currently equivalent to qwen3.5-plus-2026-02-15
International0<Token≤256K$0.4$2.4$2.41 million tokens
256K<Token≤1M$0.5$3$3
qwen3.5-plus-2026-04-20International0<Token≤256K$0.4$2.4$2.41 million tokens
256K<Token≤1M$0.5$3$3
qwen3.5-plus-2026-02-15International0<Token≤256K$0.4$2.4$2.41 million tokens
256K<Token≤1M$0.5$3$3
qwen-plus
Currently equivalent to qwen-plus-2025-12-01
International0<Token≤256K$0.4$1.2$41 million tokens
256K<Token≤1M$1.2$3.6$12
qwen-plus-latestInternational0<Token≤256K$0.4$1.2$41 million tokens
256K<Token≤1M$1.2$3.6$12
qwen-plus-2025-12-01International0<Token≤256K$0.4$1.2$41 million tokens
256K<Token≤1M$1.2$3.6$12
qwen-plus-2025-09-11International0<Token≤256K$0.4$1.2$41 million tokens
256K<Token≤1M$1.2$3.6$12
qwen-plus-2025-07-28International0<Token≤256K$0.4$1.2$41 million tokens
256K<Token≤1M$1.2$3.6$12
qwen-plus-2025-07-14InternationalNo tiered pricing$0.4$1.2$41 million tokens
qwen-plus-2025-04-28InternationalNo tiered pricing$0.4$1.2$41 million tokens
qwen-plus-2025-01-25InternationalNo tiered pricing$0.4$1.2
1 million tokens

Qwen-Flash

You are charged for input tokens and output tokens. If the model supports batch calls, the unit price for both input and output tokens is 50% of the real-time inference price. If the model supports context cache, only input tokens receive a discount. These two discounts cannot apply simultaneously.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
Model IDDeployment scopeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3.8-flash
context caching discount
International0<Token≤1M$0.15$0.471 million tokens
qwen3.7-flash
Currently equivalent to qwen3.7-flash-2026-07-15
50% batch inference discount
context caching discount
International0<Token≤32K$0.030$0.1301 million tokens
32K<Token≤256K$0.100$0.400
256K<Token≤1M$0.200$0.800
qwen3.7-flash-2026-07-15International0<Token≤32K$0.030$0.1301 million tokens
32K<Token≤256K$0.100$0.400
256K<Token≤1M$0.200$0.800
qwen3.6-flash
Currently equivalent to qwen3.6-flash-2026-04-16
50% batch inference discount
context caching discount
International0<Token≤256K$0.25$1.51 million tokens
256K<Token≤1M$1$4
qwen3.6-flash-2026-04-16International0<Token≤256K$0.25$1.51 million tokens
256K<Token≤1M$1$4
qwen3.5-flash
Currently equivalent to qwen3.5-flash-2026-02-23
50% batch inference discount
context caching discount
International0<Token≤1M$0.1$0.41 million tokens
qwen3.5-flash-2026-02-23International0<Token≤1M$0.1$0.41 million tokens
qwen-flash
Currently equivalent to qwen-flash-2025-07-28
50% batch inference discount
context caching discount
International0<Token≤256K$0.05$0.41 million tokens
256K<Token≤1M$0.25$2
qwen-flash-2025-07-28International0<Token≤256K$0.05$0.41 million tokens
256K<Token≤1M$0.25$2

Qwen-Turbo

Qwen-Turbo will no longer be updated. We recommend switching to Qwen-Flash.
You are charged for input tokens and output tokens. If the model supports batch calls, the unit price for both input and output tokens is 50% of the real-time inference price.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
Non-Thinking modeThinking mode (chain of thought + answer)
qwen-turbo
50% batch inference discount
International$0.05$0.2$0.51 million tokens

QwQ

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwq-plus

International

$0.8

$2.4

1 million tokens

Qwen-Long

You are charged for input tokens and output tokens.
  • China (Beijing)

Model ID

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

qwen-long-latest

$0.072

$0.287

No free quota

qwen-long-2025-01-25

$0.072

$0.287

No free quota

Qwen-Omni

Pricing rule: billed by input tokens and output tokens. For the token calculation rules of different modalities, see Billing and rate limits.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
Text/Image/videoAudioText
Multimodal input
Text + audio
Audio only billed
qwen3.5-omni-plus
Currently equivalent to qwen3.5-omni-plus-2026-03-15
International$1.4$11$8.3$441 million tokens
qwen3.5-omni-plus-2026-03-15International$1.4$11$8.3$441 million tokens
qwen3.5-omni-flash
Currently equivalent to qwen3.5-omni-flash-2026-03-15
International$0.4$3$2.2$11.91 million tokens
qwen3.5-omni-flash-2026-03-15International$0.4$3$2.2$11.91 million tokens
Model IDDeployment scopeModeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
TextAudioImage/videoText
Text-only input
Text
Multimodal input
Text + audio
Audio only billed
qwen3-omni-flash
Currently equivalent to qwen3-omni-flash-2025-12-01
InternationalNon-Thinking and Thinking modes$0.43$3.81$0.78$1.66$3.06$15.111 million tokens (regardless of modality)
qwen3-omni-flash-2025-12-01InternationalNon-Thinking and Thinking modes$0.43$3.81$0.78$1.66$3.06$15.111 million tokens (regardless of modality)
qwen3-omni-flash-2025-09-15InternationalNon-Thinking and Thinking modes$0.43$3.81$0.78$1.66$3.06$15.111 million tokens (regardless of modality)
qwen-omni-turbo
Currently equivalent to qwen-omni-turbo-2025-03-26
InternationalNon-Thinking mode$0.07$4.44$0.21$0.27$0.63$8.891 million tokens (regardless of modality)
qwen-omni-turbo-latestInternationalNon-Thinking mode$0.07$4.44$0.21$0.27$0.63$8.891 million tokens (regardless of modality)
qwen-omni-turbo-2025-03-26InternationalNon-Thinking mode$0.07$4.44$0.21$0.27$0.63$8.891 million tokens (regardless of modality)

Qwen-Omni-Realtime

Pricing rule: billed by input tokens and output tokens. For the token calculation rules of different modalities, see Billing and rate limits.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
Text/imageAudioText
Multimodal input
Text + audio
Audio only billed
qwen3.5-omni-plus-realtime
Currently equivalent to qwen3.5-omni-plus-realtime-2026-03-15
International$2.1$16.5$12.4$621 million tokens
qwen3.5-omni-plus-realtime-2026-03-15International$2.1$16.5$12.4$621 million tokens
qwen3.5-omni-flash-realtime
Currently equivalent to qwen3.5-omni-flash-realtime-2026-03-15
International$0.55$4.5$3.3$17.71 million tokens
qwen3.5-omni-flash-realtime-2026-03-15International$0.55$4.5$3.3$17.71 million tokens
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
TextAudioImageText
Text-only input
Text
Multimodal input
Text + audio
Audio only billed
qwen3-omni-flash-realtime
Currently equivalent to qwen3-omni-flash-realtime-2025-12-01
International$0.52$4.57$0.94$1.99$3.67$18.131 million tokens (regardless of modality)
qwen3-omni-flash-realtime-2025-12-01International$0.52$4.57$0.94$1.99$3.67$18.131 million tokens (regardless of modality)
qwen3-omni-flash-realtime-2025-09-15International$0.52$4.57$0.94$1.99$3.67$18.131 million tokens (regardless of modality)
qwen-omni-turbo-realtime
Currently equivalent to qwen-omni-turbo-realtime-2025-05-08
International$0.270$4.440$0.840$1.070$2.520$8.8901 million tokens (regardless of modality)
qwen-omni-turbo-realtime-latestInternational$0.270$4.440$0.840$1.070$2.520$8.8901 million tokens (regardless of modality)
qwen-omni-turbo-realtime-2025-05-08International$0.270$4.440$0.840$1.070$2.520$8.8901 million tokens (regardless of modality)

QVQ

Pricing rule: billed by input tokens and output tokens. For the token calculation rules of different modalities, see Billing and rate limits.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qvq-max

International

$1.2

$4.8

1 million tokens

Qwen-VL

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)
  • Germany (Frankfurt)
  • US (Virginia)
Model IDDeployment scopeModeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought + answer
Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-vl-plus
Currently equivalent to qwen3-vl-plus-2025-12-19
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤32K$0.2$1.61 million tokens
32K<Token≤128K$0.3$2.4
128K<Token≤256K$0.6$4.8
qwen3-vl-plus-2025-12-19InternationalNon-Thinking and Thinking modes0<Token≤32K$0.2$1.61 million tokens
32K<Token≤128K$0.3$2.4
128K<Token≤256K$0.6$4.8
qwen3-vl-plus-2025-09-23InternationalNon-Thinking and Thinking modes0<Token≤32K$0.2$1.61 million tokens
32K<Token≤128K$0.3$2.4
128K<Token≤256K$0.6$4.8
qwen3-vl-flash
Currently equivalent to qwen3-vl-flash-2026-01-22
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤32K$0.05$0.41 million tokens
32K<Token≤128K$0.075$0.6
128K<Token≤256K$0.12$0.96
qwen3-vl-flash-2026-01-22InternationalNon-Thinking and Thinking modes0<Token≤32K$0.05$0.41 million tokens
32K<Token≤128K$0.075$0.6
128K<Token≤256K$0.12$0.96
qwen3-vl-flash-2025-10-15InternationalNon-Thinking and Thinking modes0<Token≤32K$0.05$0.41 million tokens
32K<Token≤128K$0.075$0.6
128K<Token≤256K$0.12$0.96
More models
Model IDDeployment scopeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen-vl-max
context caching discount
InternationalNo tiered pricing$0.8$3.21 million tokens
qwen-vl-plus
context caching discount
InternationalNo tiered pricing$0.21$0.631 million tokens

Qwen-OCR

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen-vl-ocr
Currently equivalent to qwen-vl-ocr-2025-11-20
International$0.07$0.161 million tokens
qwen-vl-ocr-2025-11-20International1 million tokens

Qwen Math

You are charged for input tokens and output tokens.
  • China (Beijing)

Model ID

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

qwen-math-plus

$0.574

$1.721

No free quota

qwen-math-plus-latest

$0.574

$1.721

No free quota

qwen-math-plus-2024-09-19

$0.574

$1.721

No free quota

qwen-math-plus-2024-08-16

$0.574

$1.721

No free quota

qwen-math-turbo

$0.287

$0.861

No free quota

Qwen-Coder

You are charged for input tokens and output tokens. If the model supports context cache, only input tokens receive a discount.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
Model IDDeployment scopeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-coder-plus
Currently equivalent to qwen3-coder-plus-2025-09-23
context caching discount
International0<Token≤32K$1$51 million tokens
32K<Token≤128K$1.8$9
128K<Token≤256K$3$15
256K<Token≤1M$6$60
qwen3-coder-plus-2025-09-23International0<Token≤32K$1$51 million tokens
32K<Token≤128K$1.8$9
128K<Token≤256K$3$15
256K<Token≤1M$6$60
qwen3-coder-plus-2025-07-22International0<Token≤32K$1$51 million tokens
32K<Token≤128K$1.8$9
128K<Token≤256K$3$15
256K<Token≤1M$6$60
qwen3-coder-flash
Currently equivalent to qwen3-coder-flash-2025-07-28
International0<Token≤32K$0.3$1.51 million tokens
32K<Token≤128K$0.5$2.5
128K<Token≤256K$0.8$4
256K<Token≤1M$1.6$9.6
qwen3-coder-flash-2025-07-28International0<Token≤32K$0.3$1.51 million tokens
32K<Token≤128K$0.5$2.5
128K<Token≤256K$0.8$4
256K<Token≤1M$1.6$9.6

Qwen Translation

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-mt-plus

International

$2.46

$7.37

1 million tokens

qwen-mt-flash

International

$0.16

$0.49

1 million tokens

qwen-mt-lite

International

$0.12

$0.36

1 million tokens

qwen-mt-turbo

International

$0.16

$0.49

1 million tokens

Qwen Data Mining

You are charged for input tokens and output tokens.
  • China (Beijing)

Model ID

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

qwen-doc-turbo

$0.087

$0.144

No free quota

Qwen Deep Research

You are charged for input tokens and output tokens.
  • China (Beijing)

Model ID

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

qwen-deep-research

$7.742

$23.367

None

Text generation - Qwen (open source)

Qwen3.8

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeModeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought + answer
Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3.8-2.4t-a95b
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤1M$2$61 million tokens
qwen3.8-27b
context caching discount
InternationalNon-Thinking and Thinking modes0<Token≤1M$0.5$31 million tokens

Qwen3.6

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Input tokens per request

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

Non-Thinking mode

Thinking mode (chain of thought + answer)

qwen3.6-35b-a3b

International

0<Token≤256K

$0.375

$2.25

$2.25

1 million tokens

qwen3.6-27b

International

0<Token≤256K

$0.6

$3.6

$3.6

1 million tokens

Qwen3.5

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Input tokens per request

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

Non-Thinking mode

Thinking mode (chain of thought + answer)

qwen3.5-397b-a17b

International

0<Token≤256K

$0.6

$3.6

$3.6

1 million tokens

qwen3.5-122b-a10b

International

0<Token≤256K

$0.4

$3.2

$3.2

1 million tokens

qwen3.5-27b

International

0<Token≤256K

$0.3

$2.4

$2.4

1 million tokens

qwen3.5-35b-a3b

International

0<Token≤256K

$0.25

$2

$2

1 million tokens

Qwen3

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Mode

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

Non-Thinking mode

Thinking mode

qwen3-next-80b-a3b-thinking

International

Thinking mode only

$0.15

-

$1.2

1 million tokens

qwen3-next-80b-a3b-instruct

International

Non-Thinking mode only

$0.15

$1.2

-

1 million tokens

qwen3-235b-a22b-thinking-2507

International

Thinking mode only

$0.23

-

$2.3

1 million tokens

qwen3-235b-a22b-instruct-2507

International

Non-Thinking mode only

$0.23

$0.92

-

1 million tokens

qwen3-30b-a3b-thinking-2507

International

Thinking mode only

$0.2

-

$2.4

1 million tokens

qwen3-30b-a3b-instruct-2507

International

Non-Thinking mode only

$0.2

$0.8

-

1 million tokens

qwen3-235b-a22b

International

Non-Thinking and Thinking modes

$0.7

$2.8

$8.4

1 million tokens

qwen3-32b

International

Non-Thinking and Thinking modes

$0.16

$0.64

$0.64

1 million tokens

qwen3-30b-a3b

International

Non-Thinking and Thinking modes

$0.2

$0.8

$2.4

1 million tokens

qwen3-14b

International

Non-Thinking and Thinking modes

$0.35

$1.4

$4.2

1 million tokens

qwen3-8b

International

Non-Thinking and Thinking modes

$0.18

$0.7

$2.1

1 million tokens

Qwen-Omni

Pricing rule: billed by input tokens and output tokens. For the token calculation rules of different modalities, see Billing and rate limits.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
TextAudioImage/videoText
Text-only input
Text
Multimodal input
Text + audio
Audio only billed
qwen2.5-omni-7bInternational$0.10$6.76$0.28$0.40$0.84$13.511 million tokens (regardless of modality)

Qwen3-Omni-Captioner

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen3-omni-30b-a3b-captioner

International

$3.81

$3.06

1 million tokens

Qwen-VL

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
Model IDDeployment scopeModeInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought + answer
Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-vl-235b-a22b-thinkingInternationalThinking mode only$0.4$41 million tokens
qwen3-vl-235b-a22b-instructInternationalNon-Thinking mode only$0.4$1.61 million tokens
qwen3-vl-32b-thinkingInternationalThinking mode only$0.16$0.641 million tokens
qwen3-vl-32b-instructInternationalNon-Thinking mode only$0.16$0.641 million tokens
qwen3-vl-30b-a3b-thinkingInternationalThinking mode only$0.2$2.41 million tokens
qwen3-vl-30b-a3b-instructInternationalNon-Thinking mode only$0.2$0.81 million tokens
qwen3-vl-8b-thinkingInternationalThinking mode only$0.18$2.11 million tokens
qwen3-vl-8b-instructInternationalNon-Thinking mode only$0.18$0.71 million tokens
More models

Model ID

Deployment scope

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

Qwen-Coder

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Input tokens per request

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen3-coder-next

International

0<Token≤32K

$0.3

$1.5

1 million tokens

32K<Token≤128K

$0.5

$2.5

128K<Token≤256K

$0.8

$4

qwen3-coder-480b-a35b-instruct

International

0<Token≤32K

$1.5

$7.5

1 million tokens

32K<Token≤128K

$2.7

$13.5

128K<Token≤200K

$4.5

$22.5

qwen3-coder-30b-a3b-instruct

International

0<Token≤32K

$0.45

$2.25

1 million tokens

32K<Token≤128K

$0.75

$3.75

128K<Token≤200K

$1.2

$6

Text generation - third-party models

DeepSeek

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
DeepSeek-V4-Flash-0731 has been adjusted to peak/off-peak pricing since 2026-08-17 00:00. The adjusted prices are shown in the table below. For more information, see DeepSeek-V4-Flash-0731 price adjustment notice.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
  • China(Hongkong)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
deepseek-v4-pro
context caching discount
International$2.400$4.8001 million tokens
deepseek-v4-pro-0813
context caching discount
InternationalBusy hours: $1.32Idle hours: $0.66Busy hours: $3.96Idle hours: $1.981 million tokens
deepseek-v4-flash-0731
context caching discount
InternationalBusy hours: $0.44Idle hours: $0.22Busy hours: $1.32Idle hours: $0.661 million tokens
deepseek-v4-flash
context caching discount
International$0.200$0.4001 million tokens
deepseek-v3.2
context caching discount
International$0.57$1.711 million tokens

Kimi

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
  • Singapore
  • China(Hong Kong)

Model ID

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

kimi-k3

$2.827

$14.133

No free quota

kimi-k2.7-code

$0.894

$3.713

No free quota

kimi-k2.6

$0.8939

$3.7131

No free quota

kimi-k2.5

$0.574

$3.011

No free quota

kimi-k2-thinking

$0.574

$2.294

No free quota

Moonshot-Kimi-K2-Instruct

$0.574

$2.294

No free quota

MiniMax

You are charged for input tokens and output tokens.
  • China (Beijing)
Model IDModeInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought and answer
MiniMax-M2.5Thinking mode only$0.304$1.213

GLM

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
  • China (Hong Kong)
Model IDDeployment scopeModeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought and answer
Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
glm-5.2InternationalNon-Thinking and Thinking modesflat-rate pricing$1.400$4.400None
glm-5.2-fast-previewInternationalNon-Thinking and Thinking modesflat-rate pricing$2.800$8.800None
glm-5.1InternationalNon-Thinking and Thinking modes0<Token≤200K$1.400$4.4001 million tokens

GLM-Zhipu AI

You are charged for input tokens and output tokens.
  • Singapore
Model IDDeployment scopeModeInput tokens per requestInput price (per 1 million tokens)Output price (per 1 million tokens)
Chain of thought and answer
Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
ZHIPU/GLM-5.2InternationalNon-Thinking and Thinking modesflat-rate pricing$1.400$4.400None

Image generation

You are charged based on the number of input images and the number of successfully generated images. For models where the input image price is not specified, you are not charged for input. You are charged for output based on the number of successfully generated images. Formula: Cost = Input image unit price × Number of input images + Image unit price × Number of images generated. Notes: Failed requests incur no cost and do not consume your free quota.
Assume the output image unit price is $0.10 per image. If you call the API to generate four images but only three image URLs return successfully, the system charges only for the three successfully generated images.
  • Number billed: 3 images.
  • Cost calculation: 0.1 × 3 = $0.3.

Qwen Image Generation and Editing

Billed by the number of input and output images. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)
  • Germany (Frankfurt)
  • Japan (Tokyo)

Model ID

Service deployment scope

Output image resolution

Input unit price

Output unit price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-image-3.0-pro

International

1k

$0.003/image

$0.04/image

10 images

2k

$0.075/image

qwen-image-3.0

International

1k

$0.003/image

$0.03/image

10 images

2k

$0.03/image

Qwen Text-to-Image

Only output is billed. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeOutput priceFree quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen-image-2.0-pro
Currently equivalent to qwen-image-2.0-pro-2026-04-22
International$0.075/image100 images
qwen-image-2.0-pro-2026-06-22International$0.075/image100 images
qwen-image-2.0-pro-2026-04-22International$0.075/image100 images
qwen-image-2.0-pro-2026-03-03International$0.075/image100 images
qwen-image-2.0
Currently equivalent to qwen-image-2.0-2026-03-03
International$0.035/image100 images
qwen-image-2.0-2026-03-03International$0.035/image100 images
qwen-image-max
Currently equivalent to qwen-image-max-2025-12-30
International$0.075/image100 images
qwen-image-max-2025-12-30International$0.075/image100 images
qwen-image-plus
Currently equivalent to qwen-image
International$0.03/image100 images
qwen-image-plus-2026-01-09International$0.03/image100 images
qwen-imageInternational$0.035/image100 images

Qwen Image Editing

Only output is billed. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeOutput priceFree quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen-image-2.0-pro
Currently equivalent to qwen-image-2.0-pro-2026-04-22
International$0.075/image100 images
qwen-image-2.0-pro-2026-06-22International$0.075/image100 images
qwen-image-2.0-pro-2026-04-22International$0.075/image100 images
qwen-image-2.0-pro-2026-03-03International$0.075/image100 images
qwen-image-2.0
Currently equivalent to qwen-image-2.0-2026-03-03
International$0.035/image100 images
qwen-image-2.0-2026-03-03International$0.035/image100 images
qwen-image-edit-max
Currently equivalent to qwen-image-edit-max-2026-01-16
International$0.075/image100 images
qwen-image-edit-max-2026-01-16International$0.075/image100 images
qwen-image-edit-plus
Currently equivalent to qwen-image-edit-plus-2025-10-30
International$0.03/image100 images
qwen-image-edit-plus-2025-12-15International$0.03/image100 images
qwen-image-edit-plus-2025-10-30International$0.03/image100 images
qwen-image-editInternational$0.045/image100 images

Qwen Image Translation

Only output is billed. For pricing rules, see Image generation.
  • Singapore
  • China (Beijing)

Model ID

Service deployment scope

Output price

Free quota(Note)

qwen-mt-image-2.0

International

$0.0006/image

100 images

Qwen-Text-to-Image-Z-Image

Only output is billed. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

z-image-turbo

International

Prompt rewriting disabled (prompt_extend=false): $0.015/image

Prompt rewriting enabled (prompt_extend=true): $0.03/image

100 images

Wanx Text-to-Image

Only output is billed. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.6-t2i

International

$0.03/image

50 images

wan2.5-t2i-preview

International

$0.03/image

50 images

wan2.2-t2i-plus

International

$0.05/image

100 images

wan2.2-t2i-flash

International

$0.025/image

100 images

wan2.1-t2i-plus

International

$0.05/image

200 images

wan2.1-t2i-turbo

International

$0.025/image

200 images

Wanx Image Generation and Editing

Only output is billed. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • US (Virginia)

Model ID

Deployment scope

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.7-image-pro

International

$0.075/image

50 images

wan2.7-image

International

$0.03/image

50 images

wan2.6-image

International

$0.03/image

50 images

Wanx General Image Editing

Only output is billed. For pricing rules, see Image generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.5-i2i-preview

International

$0.03/image

50 images

AIVirtual Try-on - OutfitAnyone

  • aitryon-plus: Input is free while output is billed. For pricing rules, see Image generation.
  • aitryon-parsing-v1: Input is billed while output is free. Billed by the number of input images. Failed requests are not billed.
  • China (Beijing)

Model ID

Unit price

Free quota(Note)

aitryon-plus

$0.071677/image

No free quota

aitryon-parsing-v1

$0.000574/image

Video generation

You are not charged for input. You are charged for output based on the total duration of successfully generated videos (in seconds). Formula: Cost = Video unit price × Video duration (seconds). Notes:
  • Some models charge by output video resolution. Prices differ for resolutions such as 480P, 720P, and 1080P.
  • Some models charge by output video edition. Prices differ for editions such as Standard Edition and Professional Edition.
  • Some models charge by output video aspect ratio. Prices differ for aspect ratios such as 1:1 and 3:4.
  • Some models use a flat rate, regardless of resolution, edition, or aspect ratio.
  • Failed requests incur no cost and do not consume your free quota.

HappyHorse-Text-to-video

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)

Model ID

Deployment scope

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

happyhorse-1.1-t2v

International

480P

List price $0.07/second (Limited-time 40% off)

10 seconds

720P

List price $0.14/second (Limited-time 40% off)

1080P

List price $0.18/second (Limited-time 40% off)

happyhorse-1.0-t2v

International

720P

List price $0.14/second (Limited-time 20% off)

10 seconds

1080P

List price $0.24/second (Limited-time 20% off)

HappyHorse-Image-to-video - first frame

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)

Model ID

Deployment scope

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

happyhorse-1.1-i2v

International

480P

List price $0.07/second (Limited-time 40% off)

10 seconds

720P

List price $0.14/second (Limited-time 40% off)

1080P

List price $0.18/second (Limited-time 40% off)

happyhorse-1.0-i2v

International

720P

List price $0.14/second (Limited-time 20% off)

10 seconds

1080P

List price $0.24/second (Limited-time 20% off)

HappyHorse-Reference-to-video

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)

Model ID

Deployment scope

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

happyhorse-1.1-r2v

International

480P

List price $0.07/second (Limited-time 40% off)

10 seconds

720P

List price $0.14/second (Limited-time 40% off)

1080P

List price $0.18/second (Limited-time 40% off)

happyhorse-1.0-r2v

International

720P

List price $0.14/second (Limited-time 20% off)

10 seconds

1080P

List price $0.24/second (Limited-time 20% off)

HappyHorse-Video editing

The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
Pricing rule: both input and output videos are billed by video duration (seconds). Failed requests are not billed and do not consume the free quota.

Model ID

Deployment scope

Output video resolution

Input and output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

happyhorse-1.0-video-edit

International

720P

List price $0.14/second (Limited-time 20% off)

10 seconds

1080P

List price $0.24/second (Limited-time 20% off)

Wan 3.0-Video Generation

Pricing rule: both input and output videos are billed by video duration (seconds). Failed requests are not billed and do not consume the free quota. Billing formula: billable duration = input video duration + output video duration.
  • The billable duration of the input video is the actual input video duration in seconds.
  • The billable duration of the output video is the duration (in seconds) of successfully generated videos.
  • The free quota is 30 seconds combined for input and output video duration.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Japan (Tokyo)
  • Germany (Frankfurt)
  • US (Virginia)
  • China (Hong Kong)

Model ID

Service deployment scope

Output video resolution

Input and output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan3.0-video-prime

International

480P

$0.068/second

30 seconds

720P

$0.14/second

1080P

$0.28/second

wan3.0-video

International

480P

List price $0.05/second (Limited-time 30% off)

30 seconds

720P

List price $0.1/second (Limited-time 30% off)

1080P

List price $0.2/second (Limited-time 30% off)

Wanx-Text-to-Video

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.7-t2v-2026-06-12

International

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.7-t2v-2026-04-25

International

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.7-t2v

International

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.6-t2v

International

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.5-t2v-preview

International

480P

$0.05/second

50 seconds

720P

$0.10/second

1080P

$0.15/second

wan2.2-t2v-plus

International

480P

$0.02/second

50 seconds

1080P

$0.10/second

wan2.1-t2v-turbo

International

480P

$0.036/second

50 seconds

720P

$0.036/second

wan2.1-t2v-plus

International

720P

$0.10/second

50 seconds

Wanx-Image-to-Video

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Output video type

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.7-i2v-2026-04-25

International

Audio video

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.7-i2v

International

Audio video

720P

$0.10/second

50 seconds

1080P

$0.15/second

Wanx-Image-to-Video-First-Frame

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Output video type

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.6-i2v-flash

International

Audio video

audio=true

720P

$0.05/second

50 seconds

1080P

$0.075/second

Silent video

audio=false

720P

$0.025/second

1080P

$0.0375/second

wan2.6-i2v

International

Audio video

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.5-i2v-preview

International

Audio video

480P

$0.05/second

50 seconds

720P

$0.10/second

1080P

$0.15/second

wan2.2-i2v-flash

International

Silent video

480P

$0.015/second

50 seconds

720P

$0.036/second

wan2.2-i2v-plus

International

Silent video

480P

$0.02/second

50 seconds

1080P

$0.10/second

wan2.1-t2v-turbo

International

Silent video

480P

$0.036/second

50 seconds

720P

$0.036/second

wan2.1-t2v-plus

International

Silent video

720P

$0.10/second

50 seconds

Wanx-Image-to-Video-First-Last-Frame

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.2-kf2v-flash

International

480P

$0.015/second

50 seconds

720P

$0.036/second

1080P

$0.07/second

wan2.1-kf2v-plus

International

720P

$0.10/second

50 seconds

Wanx-Reference-to-Video

Pricing rule: both input and output videos are billed by video duration (seconds). Failed requests are not billed and do not consume the free quota. Billing formula: billable duration = input video duration (up to 5 seconds) + output video duration.
  • The billable duration of the input video does not exceed 5 seconds. For calculation rules, see Billing and rate limiting.
  • The billable duration of the output video is duration (in seconds) of successfully generated videos.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Germany (Frankfurt)
  • US (Virginia)

Model ID

Deployment scope

Output video type

Output video resolution

Input and output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.7-r2v-2026-06-12

International

Audio video

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.7-r2v

International

Audio video

720P

$0.10/second

50 seconds

1080P

$0.15/second

wan2.6-r2v-flash

International

Audio video

audio=true

720P

$0.05/second

50 seconds

1080P

$0.075/second

Silent video

audio=false

720P

$0.025/second

1080P

$0.0375/second

wan2.6-r2v

International

Audio video

720P

$0.10/second

50 seconds

1080P

$0.15/second

Wanx-Video-Editing

The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Pricing rule: both input and output videos are billed by video duration (seconds). Failed requests are not billed and do not consume the free quota.

Model ID

Deployment scope

Output video resolution

Input and output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.7-videoedit

International

720P

$0.10/second

50 seconds

1080P

$0.15/second

Pricing rule: input is free. Output video is billed by video duration (seconds). Failed requests are not billed and do not consume the free quota.

Model ID

Deployment scope

Output video resolution

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.1-vace-plus

International

720P

$0.10/second

50 seconds

Wanx-Digital Human

  • wan2.2-s2v-detect: Input is billed while output is free. Input is billed by the number of images processed. Each input image is billed once as long as the request succeeds, regardless of the detection result.
  • wan2.2-s2v: Input is free while output is billed. Output is billed by the duration (in seconds) of successfully generated videos. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Unit price

Free quota(Note)

wan2.2-s2v-detect

Input image: $0.000574/image

No free quota

wan2.2-s2v

Output video:

  • 480P: $0.071677/second

  • 720P: $0.129018/second

No free quota

Wanx-Image-to-Motion

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Output video mode

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.2-animate-move

International

Standard modewan-std

$0.12/second

50 seconds

Professional modewan-pro

$0.18/second

Wanx-Video-Face-Swap

Only output is billed. For pricing rules, see Video generation.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Output video mode

Output price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

wan2.2-animate-mix

International

Standard modewan-std

$0.18/second

50 seconds

Professional modewan-pro

$0.26/second

AnimateAnyone

  • animate-anyone-detect-gen2: Input is billed while output is free. Input is billed by the number of images processed. Each input image is billed once as long as the request succeeds, regardless of the detection result.
  • animate-anyone-template-gen2: Input is free while output is billed. Output is billed by the duration (in seconds) of successfully generated videos. For pricing rules, see Video generation.
  • animate-anyone-gen2: Input is free while output is billed. Output is billed by the duration (in seconds) of successfully generated videos. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Unit price

Free quota(Note)

animate-anyone-detect-gen2

Input image: $0.000574/image

No free quota

animate-anyone-template-gen2

Output video: $0.011469/second

No free quota

animate-anyone-gen2

Output video: $0.011469/second

No free quota

EMO

  • emo-detect-v1: Input is billed while output is free. Input is billed by the number of images processed. Each input image is billed once as long as the request succeeds, regardless of the detection result.
  • emo-v1: Input is free while output is billed. Output is billed by the duration (in seconds) of successfully generated videos. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Unit price

Free quota(Note)

emo-detect-v1

Input image: $0.000574/image

No free quota

emo-v1

Output video:

  • 1:1landscape video: $0.011469/second

  • 3:4landscape video: $0.022937/second

LivePortrait

  • liveportrait-detect: Input is billed while output is free. Input is billed by the number of images processed. Each input image is billed once as long as the request succeeds, regardless of the detection result.
  • liveportrait: Input is free while output is billed. Output is billed by the duration (in seconds) of successfully generated videos. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Unit price

Free quota(Note)

liveportrait-detect

Input image: $0.000574/image

No free quota

liveportrait

Output video: $0.002868/second

Emoji Sticker

  • emoji-detect-v1: Input is billed while output is free. Input is billed by the number of images processed. Each input image is billed once as long as the request succeeds, regardless of the detection result.
  • emoji-v1: Input is free while output is billed. Output is billed by the duration (in seconds) of successfully generated videos. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Unit price

Free quota(Note)

emoji-detect-v1

Input image: $0.000574/image

No free quota

emoji-v1

Output video: $0.011469/second

VideoRetalk

Only output is billed. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Output price

Free quota(Note)

videoretalk

$0.011469/second

No free quota

Video Style Repaint

Only output is billed. For pricing rules, see Video generation.
  • China (Beijing)

Model ID

Output video resolution

Output price

Free quota(Note)

video-style-transform

540P

$0.028671/second

No free quota

720P

$0.071677/second

Music generation

Pricing rule: billed by the duration (in seconds) of output audio. Input is free.
  • China (Beijing)

Model ID

Output price (per second)

Free quota(Note)

fun-music-preview

$0.000695

No free quota

fun-music-v1

$0.000275

Speech synthesis (text-to-speech)

Character counting rules: For the models in this section that are billed by character count (input prices are listed per 10,000 characters), the number of characters in the input text is counted as follows:
  • Each Chinese character (including simplified Chinese characters, traditional Chinese characters, Japanese kanji, and Korean hanja) counts as 2 characters.
  • Each other character (such as an English letter, a digit, a punctuation mark, a space, a Japanese kana, or a Korean letter) counts as 1 character.
  • When SSML is used, the SSML tags themselves are not counted. Only the text content to be synthesized is counted.
Examples: “你好” is 4 characters (2+2); “中A文123” is 8 characters (2+1+2+1+1+1); “中文。” is 5 characters (2+2+1); “中 文。” is 6 characters (2+1+2+1).

Qwen-Audio-TTS

Billing rules: Fees are charged based on the number of characters in the input text. Output is not billed.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Service deployment scope

Input unit price (per 10,000 characters)

Free quota(note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-audio-3.0-tts-plus

International

$0.2

10,000 characters

qwen-audio-3.0-tts-flash

International

$0.15

10,000 characters

Qwen-TTS

The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Qwen3-TTS-Instruct-Flash
  • Qwen3-TTS-VD
  • Qwen3-TTS-VC
  • Qwen3-TTS-Flash
Pricing rule: billed by the number of input text characters. Output is free.
Model IDDeployment scopeInput price (per 10,000 characters)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-tts-instruct-flash
Currently equivalent to qwen3-tts-instruct-flash-2026-01-26
International$0.115110,000 characters
qwen3-tts-instruct-flash-2026-01-26International$0.115110,000 characters

Qwen-TTS-Realtime

The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Qwen3-TTS-Instruct-Flash-Realtime
  • Qwen3-TTS-VD-Realtime
  • Qwen3-TTS-VC-Realtime
  • Qwen3-TTS-Flash-Realtime
Pricing rule: billed by the number of input text characters. Output is free.
Model IDDeployment scopeInput price (per 10,000 characters)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-tts-instruct-flash-realtime
Currently equivalent to qwen3-tts-instruct-flash-realtime-2026-01-22
International$0.143110,000 characters
qwen3-tts-instruct-flash-realtime-2026-01-22International$0.143110,000 characters

Qwen-TTS Voice cloning

Pricing rule: billed by the number of new voice clones created.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Price (per voice clone)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-voice-enrollment

International

$0.01

1,000 voices/account

Qwen-TTS Voice design

Pricing rule: billed by the number of new voice clones created.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Price (per voice clone)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-voice-design

International

$0.2

10 voices/account

CosyVoice

The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Pricing rule: billed by the number of input text characters. Output is free.

Model ID

Deployment scope

Input price (per 10,000 characters)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

cosyvoice-v3-plus

International

$0.26

10,000 characters

cosyvoice-v3-flash

International

$0.13

10,000 characters

Speech recognition (speech-to-text) and translation (speech-to-text in a specified language)

Qwen-LiveTranslate-Flash-Realtime

Pricing rule: billed by input tokens and output tokens. For the token calculation rules of different modalities, see Billing.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
Input: audioInput: imageOutput: textOutput: audio
qwen3.5-livetranslate-flash-realtimeInternational$7.5$0.55$20$301 million tokens
qwen3.5-livetranslate-flash-realtime-2026-05-19International$7.5$0.55$20$301 million tokens
qwen3-livetranslate-flash-realtime
Currently equivalent to qwen3-livetranslate-flash-realtime-2025-09-22
International$10$1.3$10$381 million tokens
qwen3-livetranslate-flash-realtime-2025-09-22International$10$1.3$10$381 million tokens

Qwen-LiveTranslate-Flash

Pricing rule: billed by input tokens and output tokens. For the token calculation rules of different modalities, see Billing.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

Input: audio

Input: image

Output: text

Output: audio

qwen3-livetranslate-flash

International

$1.577

$0.631

$1.577

$6.308

1 million tokens

qwen3-livetranslate-flash-2025-12-01

International

$1.577

$0.631

$1.577

$6.308

1 million tokens

Qwen-Audio-3.0-ASR-Flash-Streaming

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-audio-3.0-asr-flash-streaming

International

$0.00009/second

36,000 seconds (10 hours)

Qwen-Audio-3.0-ASR-Flash-Filetrans

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-audio-3.0-asr-flash-filetrans

International

$0.000035/second

36,000 seconds (10 hours)

Qwen-Audio-3.0-ASR-Flash

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen-audio-3.0-asr-flash

International

$0.000035/second

36,000 seconds (10 hours)

Qwen-ASR

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • US (Virginia)
Model IDDeployment scopeInput priceFree quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-asr-flash-filetransInternational$0.000035/second36,000 seconds (10 hours)
qwen3-asr-flash-filetrans-2025-11-17International$0.000035/second36,000 seconds (10 hours)
qwen3-asr-flash
Currently equivalent to qwen3-asr-flash-2025-09-08
International$0.000035/second36,000 seconds (10 hours)
qwen3-asr-flash-2026-02-10International$0.000035/second36,000 seconds (10 hours)
qwen3-asr-flash-2025-09-08International$0.000035/second36,000 seconds (10 hours)

Qwen-ASR-Realtime

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput priceFree quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen3-asr-flash-realtime
Currently equivalent to qwen3-asr-flash-realtime-2025-10-27
International$0.000090/second36,000 seconds (10 hours)
qwen3-asr-flash-realtime-2026-02-10International$0.000090/second36,000 seconds (10 hours)
qwen3-asr-flash-realtime-2025-10-27International$0.000090/second36,000 seconds (10 hours)

Fun-ASR

Audio file recognition

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
Model IDDeployment scopeInput priceFree quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
fun-asr
Currently equivalent to fun-asr-2025-11-07
International$0.000035/second36,000 seconds (10 hours)
fun-asr-2025-11-07International$0.000035/second36,000 seconds (10 hours)
fun-asr-2025-08-25International$0.000035/second36,000 seconds (10 hours)
fun-asr-mtlInternational$0.000035/second36,000 seconds (10 hours)
fun-asr-mtl-2025-08-25International$0.000035/second36,000 seconds (10 hours)
fun-asr-flash-2026-06-15International$0.000035/second36,000 seconds (10 hours)

Real-time speech recognition

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

fun-asr-realtime

International

$0.00009/second

36,000 seconds (10 hours)

fun-asr-realtime-2025-11-07

International

$0.00009/second

36,000 seconds (10 hours)

Paraformer

Audio file recognition

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
  • China (Beijing)

Model ID

Input price

paraformer-v2

$0.000012/second

paraformer-8k-v2

$0.000012/second

Real-time speech recognition

Pricing rule: billed by the duration (in seconds) of input audio. Output is free.
  • China (Beijing)

Model ID

Input price

Free quota(Note)

paraformer-realtime-v2

$0.000035/second

No free quota

paraformer-realtime-8k-v2

$0.000035/second

Voice Chat

Real-time Voice Chat

Real-time voice chat models support both text and audio input and output, billed separately by input tokens and output tokens. Audio tokens are calculated based on duration: Total tokens = Audio duration (seconds) × 12.5. Durations less than 1 second are rounded up to 1 second. In multi-turn conversations, similar to text-based LLMs, the model maintains a complete conversation context to ensure coherent dialogue. Historical conversation content is processed and billed as input for subsequent turns, so the input token count increases progressively with each turn. The specific billing rules for each content type are as follows:
  • User input audio and text: Counted as context and billed as input in each subsequent turn. Audio is billed as audio tokens, and text is billed as text tokens.
  • User-configured instructions: Billed as text tokens once per turn.
  • Model output text: Counted as context using text tokens and billed as input in each subsequent turn.
  • Model output audio: Billed as audio tokens only once at output time and not counted as context.
As the number of conversation turns increases, the accumulated context tokens grow progressively. We recommend controlling the number of turns in a single session or starting a new session at appropriate times to optimize costs.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)

Model ID

Deployment region

Input price (per million tokens)

Output price (per million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

Text

Audio

Text

Audio

qwen-audio-3.0-realtime-plus

International

$0.8

$6.4

$6.4

$24

1,000,000 tokens

qwen-audio-3.0-realtime-flash

International

$0.45

$4.5

$4.5

$15

1,000,000 tokens

Text embedding

Pricing rule: billed by input tokens. Output is free.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)

Model ID

Deployment scope

Input price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

text-embedding-v4

International

$0.07

1 million tokens

text-embedding-v3

International

$0.07

500,000 tokens

Multimodal embedding

Pricing rule: billed by input tokens. Output is free.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price (per million input tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

tongyi-embedding-vision-plus

International

$0.09

1 million tokens

tongyi-embedding-vision-flash

International

Image/video: $0.03

Text: $0.09

1 million tokens

Text reranking

Pricing rule: billed by input tokens. Output is free.
  • Singapore
  • China (Beijing)

Model ID

Deployment scope

Input price (per 1 million tokens)

Free quota(Note)

Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later

qwen3-rerank

International

$0.1

1 million tokens

Industry models

Intent understanding

  • China (Beijing)

Model ID

Input price (per 1 million tokens)

Output price (per 1 million tokens)

Free quota(Note)

tongyi-intent-detect-v3

$0.058

$0.144

No free quota

Role play

You are charged for input tokens and output tokens.
The following models offer a free quota only in Singapore. No free quota is available in other regions.
  • Singapore
  • China (Beijing)
  • Hong Kong (China)
  • Germany (Frankfurt)
  • US (Virginia)
  • Japan (Tokyo)
Model IDDeployment scopeInput price (per 1 million tokens)Output price (per 1 million tokens)Free quota(Note)Valid for 90 days from the date of Model Studio activation, model release, or application approval, whichever is later
qwen-plus-character
Session Cache discount
International$0.5$1.41 million tokens
qwen-flash-character
Session Cache discount
International$0.05$0.41 million tokens
qwen-plus-character-jaInternational$0.5$1.41 million tokens

Token consumption and cost control

The following items explain token consumption in common scenarios and how to reduce model call costs.
  • Billing for file reads by URL: When a model reads a file through a URL, the transfer of the file itself does not consume tokens. However, after the file content is parsed, the parsed text is converted into input tokens and billed accordingly. The number of tokens consumed depends on the length of the parsed text, not the size of the original file. For large files, consider using knowledge base chunking and indexing to reduce token consumption instead of passing the full file content directly.
  • Troubleshooting unusually high Credits consumption: Common causes of an unexpectedly high Credits cost for a single request include an oversized input token count and the use of Agent mode, which adds overhead from a large system prompt, tool and function definitions, and thinking-mode content. To reduce cost, try compressing historical messages, starting a new conversation, turning off thinking mode, or switching to a lightweight model (for example, a model in the Flash series).
  • Recommended models for low-frequency calls: For low-frequency call scenarios such as heartbeat detection, use a low-cost model, such as one in the Qwen Flash series.
  • Deprecated models: Models in the Qwen2.5 series and other models marked as deprecated in this document are no longer available for calling, and their pricing information can no longer be queried. Migrate to the corresponding current-generation models.

Error codes

If a model call fails and returns an error message, see Error codes for resolution.
Token Plan
Model Playground
  • Music generation
Statistics and Monitoring
Support