qwen
Qwen is the family of large language models developed by Alibaba Cloud, spanning the Qwen3.x flagship, the code-specialized Coder, the vision-focused VL, and the ultra-fast, economical Turbo tiers. Qwen3.8 supports ~1M context and is known for its strong reasoning, multimodal capabilities, and open-source ecosystem, making it one of the most widely used large models in China.
~1M token context window (Qwen3.8); visual understanding (Qwen3 VL Plus); code generation (Qwen3 Coder Plus/Flash); image generation (Qwen Image 2.0 Pro); Function Calling; JSON mode; OpenAI-compatible interface; open-source ecosystem (the Qwen series is open source).
Chinese-language reasoning and conversation; code generation and completion; image generation and editing; long-document analysis; agent workflows; enterprise knowledge bases; vector-based retrieval; multimodal Q&A.
Some models are available only in China; image generation is billed per image; no on-premises deployment (except open-source versions); the enterprise edition requires an Alibaba Cloud account; open-source versions differ in performance from the commercial API.
| Model | Input / 1M | Output / 1M | Cache Read / 1M | Per 1M |
|---|---|---|---|---|
| - | 1M | |||
| 1M | ||||
| 1M | ||||
| 1M | ||||
| 1M | ||||
| $7.15/M |
| $0.1192/M |
| 1M |
Qwen3.6 Flash qwen3.6-flash | $0.72/M | $4.29/M | $0.0715/M | 1M |
Qwen3.5 Flash qwen3.5-flash | $0.18/M | $1.79/M | $0.0075/M | 1M |
Qwen3.5 Plus qwen3.5-plus | $0.6/M | $3.58/M | $0.0596/M | 1M |
Qwen3 235B qwen3-235b | $0.3/M | $1.19/M | $0.0149/M | 1M |
Qwen3 Coder 480B qwen3-coder-480b | $2.24/M | $8.94/M | $0.0298/M | 1M |
Qwen3 VL 235B qwen3-vl-235b | $0.3/M | $1.19/M | - | 1M |
Qwen3 Coder Plus qwen3-coder-plus | $2.98/M | $29.81/M | $0.0298/M | 1M |
Qwen3 VL Plus qwen3-vl-plus | $0.45/M | $4.47/M | - | 1M |
Qwen3 Coder Flash qwen3-coder-flash | $0.75/M | $3.73/M | $0.0075/M | 1M |
Qwen Audio 3.0 TTS Plus qwen-audio-3.0-tts-plus | — | Default$0.0209/1K chars | — | 1M |
Qwen Image 3.0 Pro qwen-image-3.0-pro | — | 1K Resolution$0.0373/image 2K Resolution$0.0746/image | — | 1M |
Qwen Image 3.0 qwen-image-3.0 | — | 1K Resolution$0.0269/image 2K Resolution$0.0269/image | — | 1M |
Qwen3 ASR Flash qwen3-asr-flash | — | Default$0/sec | — | 1M |
Qwen Audio 3.0 ASR Flash qwen-audio-3.0-asr-flash | — | Default$0/sec | — | 1M |
Qwen3.7 Text Embedding qwen3.7-text-embedding | $0.07/M | $0/M | - | 1M |
Qwen3.8 2.4T A95B qwen3.8-2.4t-a95b | $1.79/M | $5.37/M | - | 1M |
Qwen3.8 27B qwen3.8-27b | $0.45/M | $1.79/M | - | 1M |
Qwen Long qwen-long | $0.07/M | $0.3/M | - | 1M |
Qwen Deep Research qwen-deep-research | $8.05/M | $24.29/M | - | 1M |
Qwen Audio 3.0 TTS Flash qwen-audio-3.0-tts-flash | — | Default$0.0149/1K chars | — | 1M |
Qwen3.5 OCR qwen3.5-ocr | $0.07/M | $0.3/M | - | 1M |
Qwen Image 2.0 Pro qwen-image-2.0-pro | — | Standard$0.0746/image | — | 1M |
Qwen qwen3-tts-instruct-flash | — | Default$0.012/1K chars | — | 1M |
Qwen qwen3-tts-flash | — | Default$0.012/1K chars | — | 1M |
Qwen3 Reranker 8B qwen3-reranker-8b | $0.02/M | $0/M | - | 1M |
Qwen Turbo qwen-turbo | $0.04/M | $0.09/M | $0.0089/M | 1M |
Qwen3 Embedding 8B qwen3-embedding-8b | $0.1/M | $0/M | - | 1M |