Qwen3 Model Family

Explore 28 curated models in the Qwen3 family and compare context, API pricing, modalities and capabilities.

28models tracked

Model directory

28 models in this view

Compare
Active

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

Context
1M
Input
$0.26
Output
$0.78
Text
View model

Over the past three months, we have continued to scale the **thinking capability** of Qwen3-30B-A3B, improving both the **quality and depth** of reasoning. We are pleased to introduce **Qwen3-30B-A3B-Thinking-2507**, featuring the following key enhancements:

Context
82K
Input
$0.20
Output
$2.4
Text
View model
Active

The Qwen3-ASR family includes Qwen3-ASR-1.7B and Qwen3-ASR-0.6B, which support language identification and ASR for 52 languages and dialects. Both leverage large-scale speech training data and the strong audio understanding capability of their foundation model, Qwen3-Omni. Experiments show that the 1.7B version achieves state-of-the-art performance among open-source ASR models and is competitive with the strongest proprietary commercial APIs. Here are the main features:

Context
0
Input
$3.33
Output
$0.00
Audio
View model
Active

The Qwen3-ASR family includes Qwen3-ASR-1.7B and Qwen3-ASR-0.6B, which support language identification and ASR for 52 languages and dialects. Both leverage large-scale speech training data and the strong audio understanding capability of their foundation model, Qwen3-Omni. Experiments show that the 1.7B version achieves state-of-the-art performance among open-source ASR models and is competitive with the strongest proprietary commercial APIs. Here are the main features:

Context
0
Input
$7.5
Output
$0.00
Audio
View model
Active

Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages —…

Context
0
Input
$35
Output
$0.00
Audio
View model
Active

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling…

Context
1M
Input
$0.195
Output
$0.975
Text
View model
Active

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and…

Context
1M
Input
$0.65
Output
$3.25
Text
View model
Active

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It…

Context
262K
Input
$0.78
Output
$3.9
Text
View model

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

Context
262K
Input
$0.10
Output
$1.1
Text
View model

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

Context
131K
Input
$0.15
Output
$1.2
Text
View model
Active

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG…

Context
41K
Input
$0.00
Output
$0.00
Text
View model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Context
131K
Input
$0.21
Output
$1.9
TextImage
View model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Context
131K
Input
$0.40
Output
$4
TextImage
View model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Context
131K
Input
$0.13
Output
$0.52
TextImage
View model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Context
131K
Input
$0.20
Output
$2.4
TextImage
View model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Context
131K
Input
$0.117
Output
$0.455
ImageText
View model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Context
131K
Input
$0.18
Output
$2.1
ImageText
View model
Active

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Context
262K
Input
$0.39
Output
$2.34
TextImageVideo
View model

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of…

Context
1M
Input
$0.26
Output
$1.56
TextImageVideo
View model

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This…

Context
1M
Input
$0.30
Output
$1.8
TextImageVideo
View model

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Context
262K
Input
$0.26
Output
$2.08
TextImageVideo
View model
Active

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Context
262K
Input
$0.195
Output
$1.56
TextImageVideo
View model
Active

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Context
262K
Input
$0.225
Output
$1.8
TextImageVideo
View model
Active

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Context
262K
Input
$0.10
Output
$0.15
TextImageVideo
View model
Active

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the…

Context
1M
Input
$0.065
Output
$0.26
TextImageVideo
View model
Active

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Context
262K
Input
$0.60
Output
$3.6
TextImageVideo
View model
Active

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in…

Context
1M
Input
$0.1875
Output
$1.13
TextImageVideo
View model
Active

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers…

Context
1M
Input
$0.325
Output
$1.95
TextImageVideo
View model
OUR METHOD

How AIToolly handles model data

Model-native facts and provider endpoint facts stay separate. Third-party catalogs support discovery and provider snapshots, while official documentation and verified model cards take priority for model facts.

01Sources and verification
02Verified
03Capability