Семейство моделей Qwen3

Изучите 28 моделей семейства Qwen3 и сравните контекст, цены и возможности.

28моделей в каталоге

Каталог моделей

В этом виде: 28

Сравнить
Активна

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

Контекст
1M
Вход
0,26 $
Выход
0,78 $
Текст
Открыть модель
Активна

Over the past three months, we have continued to scale the **thinking capability** of Qwen3-30B-A3B, improving both the **quality and depth** of reasoning. We are pleased to introduce **Qwen3-30B-A3B-Thinking-2507**, featuring the following key enhancements:

Контекст
82K
Вход
0,20 $
Выход
2,4 $
Текст
Открыть модель
Активна

The Qwen3-ASR family includes Qwen3-ASR-1.7B and Qwen3-ASR-0.6B, which support language identification and ASR for 52 languages and dialects. Both leverage large-scale speech training data and the strong audio understanding capability of their foundation model, Qwen3-Omni. Experiments show that the 1.7B version achieves state-of-the-art performance among open-source ASR models and is competitive with the strongest proprietary commercial APIs. Here are the main features:

Контекст
0
Вход
3,33 $
Выход
0,00 $
Аудио
Открыть модель
Активна

The Qwen3-ASR family includes Qwen3-ASR-1.7B and Qwen3-ASR-0.6B, which support language identification and ASR for 52 languages and dialects. Both leverage large-scale speech training data and the strong audio understanding capability of their foundation model, Qwen3-Omni. Experiments show that the 1.7B version achieves state-of-the-art performance among open-source ASR models and is competitive with the strongest proprietary commercial APIs. Here are the main features:

Контекст
0
Вход
7,5 $
Выход
0,00 $
Аудио
Открыть модель
Активна

Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages —…

Контекст
0
Вход
35 $
Выход
0,00 $
Аудио
Открыть модель
Активна

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling…

Контекст
1M
Вход
0,195 $
Выход
0,975 $
Текст
Открыть модель
Активна

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and…

Контекст
1M
Вход
0,65 $
Выход
3,25 $
Текст
Открыть модель
Активна

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It…

Контекст
262K
Вход
0,78 $
Выход
3,9 $
Текст
Открыть модель
Активна

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

Контекст
262K
Вход
0,10 $
Выход
1,1 $
Текст
Открыть модель
Активна

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

Контекст
131K
Вход
0,15 $
Выход
1,2 $
Текст
Открыть модель
Активна

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG…

Контекст
41K
Вход
0,00 $
Выход
0,00 $
Текст
Открыть модель
Активна

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Контекст
131K
Вход
0,21 $
Выход
1,9 $
ТекстИзображение
Открыть модель
Активна

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Контекст
131K
Вход
0,40 $
Выход
4 $
ТекстИзображение
Открыть модель
Активна

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Контекст
131K
Вход
0,13 $
Выход
0,52 $
ТекстИзображение
Открыть модель
Активна

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Контекст
131K
Вход
0,20 $
Выход
2,4 $
ТекстИзображение
Открыть модель
Активна

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Контекст
131K
Вход
0,117 $
Выход
0,455 $
ИзображениеТекст
Открыть модель
Активна

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Контекст
131K
Вход
0,18 $
Выход
2,1 $
ИзображениеТекст
Открыть модель
Активна

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Контекст
262K
Вход
0,39 $
Выход
2,34 $
ТекстИзображениеВидео
Открыть модель
Активна

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of…

Контекст
1M
Вход
0,26 $
Выход
1,56 $
ТекстИзображениеВидео
Открыть модель
Активна

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This…

Контекст
1M
Вход
0,30 $
Выход
1,8 $
ТекстИзображениеВидео
Открыть модель
Активна

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Контекст
262K
Вход
0,26 $
Выход
2,08 $
ТекстИзображениеВидео
Открыть модель
Активна

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Контекст
262K
Вход
0,195 $
Выход
1,56 $
ТекстИзображениеВидео
Открыть модель
Активна

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Контекст
262K
Вход
0,225 $
Выход
1,8 $
ТекстИзображениеВидео
Открыть модель
Активна

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Контекст
262K
Вход
0,10 $
Выход
0,15 $
ТекстИзображениеВидео
Открыть модель
Активна

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the…

Контекст
1M
Вход
0,065 $
Выход
0,26 $
ТекстИзображениеВидео
Открыть модель
Активна

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Контекст
262K
Вход
0,60 $
Выход
3,6 $
ТекстИзображениеВидео
Открыть модель
Активна

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in…

Контекст
1M
Вход
0,1875 $
Выход
1,13 $
ТекстИзображениеВидео
Открыть модель
Активна

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers…

Контекст
1M
Вход
0,325 $
Выход
1,95 $
ТекстИзображениеВидео
Открыть модель
OUR METHOD

Как AIToolly работает с данными моделей

Данные модели и endpoint поставщика хранятся раздельно. Сторонние каталоги используются для поиска и снимков, а официальная документация и проверенные карточки моделей имеют приоритет для данных модели.

01Источники и проверка
02Проверено
03Возможность