Skip to main content
Text Generation

qwen3-235b-a22b-thinking-2507

Open-source Qwen3 thinking model; compared to the previous version (Qwen3-235B-A22B) shows major improvements in logical ability, general capabilities, knowledge enhancement, and creativity, suitable for high-difficulty, strong-thinking scenarios.

Inference Service Provider

The inference service provider for qwen3-235b-a22b-thinking-2507 is Alibaba Cloud Model Studio.

Model Capabilities

  • China (Beijing)
  • Singapore
  • Germany (Frankfurt)
  • US (Virginia)
CapabilitySupportCapabilitySupport

Input Modality

Text

Output Modality

Text

Model Experience

Supported

Function Calling

Unsupported

Structured Outputs

Unsupported

Web Search

Unsupported

Prefix Completion

Unsupported

Context Caching

Unsupported

Batch Inference

Unsupported

Fine-tuning

Unsupported

Context Limits

ParameterValueParameterValue

Max Input Length

Max Output Length

Context Window

131072

Max Input Length (Thinking Mode)

126976

Max Output Length (Thinking Mode)

32768

Max Chain-of-Thought Length

81920

Pricing

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit Model Studio Console for promotional offers.
  • China (Beijing)
  • Singapore
  • Germany (Frankfurt)
  • US (Virginia)
Billing ItemPrice (USD)Unit

Input(Thinking)

0.287

Per 1M tokens

Output(Thinking)

2.868

Per 1M tokens

Rate Limits

  • China (Beijing)
  • Singapore
  • Germany (Frankfurt)
  • US (Virginia)
ParameterValue

RPM (Requests Per Minute)

600

TPM (Tokens Per Minute)

1,000,000

Token Plan
Model Playground
  • Music generation
Statistics and Monitoring