Skip to main content
Get Started

Recommended models

Alibaba Cloud Model Studio offers Qwen and third-party models for text, image, audio, and video.

Text generation

qwen3.8-maxqwen3.7-plusqwen3.8-flash
deepseek-v4-prodeepseek-v4-flashkimi-k2.7-code
glm-5.2ZHIPU/GLM-5.3kimi-k3
MiniMax-M2.5
More

Image & video

Understanding

Extract text descriptions or structured data from images and videos
qwen3.8-maxqwen3.7-plusqwen3.5-omni-plus
kimi-k2.7-code
More

Generation

Generate images and videos from text or images, with support for editing, reference, and high-resolution output
qwen-image-3.0-prowan2.7-image-prohappyhorse-1.1-t2v
happyhorse-1.1-i2vhappyhorse-1.1-r2vhappyhorse-1.0-video-edit
wan3.0-video
More

Audio & speech

Text-to-speech

For audiobook reading, voice broadcasting, virtual avatars, and more
qwen-audio-3.0-tts-plus
More

Speech recognition

Dedicated ASR and LLM-based approaches — choose based on accuracy and flexibility
qwen-audio-3.0-asr-flash-streamingqwen-audio-3.0-asr-flash-filetransqwen3.5-omni-plus-realtime
qwen3.5-omni-plus
More

Speech-to-speech

End-to-end voice conversation without separate ASR and TTS calls
qwen-audio-3.0-realtime-plusqwen3.5-omni-plus
More

Omni

Integrates understanding and generation capabilities across text, image, audio, and video modalities
qwen3.5-omni-plus-realtimeqwen3.5-omni-plus
More

Embeddings & reranking

Convert text or multimodal content into vectors, combined with reranking to improve retrieval accuracy
text-embedding-v4qwen3.7-text-embeddingtongyi-embedding-vision-plus
qwen3-rerank
More

View all models

Go to Model Plaza to browse all Qwen, third-party, domain-specific, and legacy models.
Token Plan
Model Playground
  • Music generation
Statistics and Monitoring
Support