Reasoning Tools JSON

2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding

alibaba/qwen3.8-max-0902 2026-09-02 1M context $1.71/M input $5.14/M output
7 providers
Reasoning Tools JSON

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.8-flash 2026-08-26 1M context $0.15/M input $0.47/M output
21 providers
Tools JSON

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

alibaba/qwen3.8-max 2026-08-03 1M context $2/M input $6/M output
28 providers
Reasoning Tools

Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows

alibaba/qwen3.8-max-preview 2026-07-19 1M context $2/M input $6/M output
6 providers
Reasoning Tools JSON

Lightweight multimodal Qwen model for high-throughput text, image, and video tasks

alibaba/qwen3.7-flash 2026-07-15 1M context $0.028/M input $0.113/M output
11 providers
Reasoning Tools

Multimodal Qwen workhorse for long-context agents, visual inputs, and coding

alibaba/qwen3.7-plus 2026-06-02 1M context $0.5/M input $3/M output
29 providers
Reasoning Tools

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.6-flash 2026-04-27 1M context $0.188/M input $1.125/M output
22 providers

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen3.6-max-preview 2026-04-20 262.144K context $1.3/M input $7.8/M output
12 providers
Reasoning Tools

Earlier Qwen multimodal workhorse for million-token agent and document tasks

alibaba/qwen3.6-plus 2026-04-02 1M context $0.5/M input $3/M output
23 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-flash 2026-02-23 1M context $0.029/M input $0.287/M output
7 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.5-plus 2026-02-16 1M context $0.4/M input $2.4/M output
13 providers

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3-vl-plus 2025-09-23 262.144K context $0.2/M input $1.6/M output
6 providers
Tools

Efficient Qwen model for fast chat, extraction, and high-volume workloads

alibaba/qwen-flash 2025-07-28 1M context $0.05/M input $0.4/M output
7 providers

Qwen coding model for software agents, repository edits, and code reasoning

alibaba/qwen3-coder-flash 2025-07-28 1M context $0.3/M input $1.5/M output
12 providers

Hosted Qwen coder for software agents, repo edits, and long-context code

alibaba/qwen3-coder-plus 2025-07-23 1.04858M context $1/M input $5/M output
14 providers
Reasoning Tools

Qwen reasoning model for deliberate problem solving, math, and coding

alibaba/qwq-plus 2025-03-05 131.072K context $0.8/M input $2.4/M output
4 providers

Qwen omni model for text, vision, audio, and multimodal agent tasks

alibaba/qwen-omni-turbo 2025-01-19 32.768K context $0.07/M input $0.27/M output
4 providers

Efficient Qwen model for fast chat, extraction, and high-volume workloads

alibaba/qwen-turbo 2024-11-01 1M context $0.05/M input $0.2/M output
5 providers
Tools

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen-vl-max 2024-04-08 131.072K context $0.8/M input $3.2/M output
5 providers
Tools

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen-max 2024-04-03 32.768K context $1.6/M input $6.4/M output
6 providers
Reasoning

Qwen instruction model for multilingual chat, reasoning, and tool use

alibaba/qwen-plus 2024-01-25 1M context $0.4/M input $1.2/M output
10 providers
Tools

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen-vl-plus 2024-01-25 131.072K context $0.21/M input $0.63/M output
4 providers

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

qwen/qwen3-vl-8b-thinking 131.072K context $0.18/M input $2.1/M output

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

qwen/qwen3-235b-a22b-thinking-2507 131.072K context $0.23/M input $2.3/M output

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

qwen/qwen3-coder-30b-a3b-instruct 262.144K context $0.07/M input $0.28/M output

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

qwen/qwen3-235b-a22b-2507 262.144K context $0.22/M input $0.88/M output

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

qwen/qwen3-30b-a3b 40.96K context $0.12/M input $0.5/M output

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

qwen/qwen3-8b 131.072K context $0.117/M input $0.455/M output

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-32b 40.96K context $0.08/M input $0.28/M output

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...

qwen/qwen-2.5-coder-32b-instruct 32.768K context $0.66/M input $1/M output

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-7b-instruct 32.768K context $0.1/M input $0.2/M output

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....

qwen/qwen3-vl-235b-a22b-thinking 131.072K context $0.4/M input $4/M output

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

qwen/qwen2.5-vl-72b-instruct 128K context $0.8/M input $1/M output

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

qwen/qwen3-next-80b-a3b-instruct:free 262.144K context Free input Free output

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

qwen/qwen3.8-max-0902 1M context $2/M input $6/M output

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28:thinking 1M context $0.26/M input $0.78/M output

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

qwen/qwen3.5-9b:batch 262.144K context $0.17/M input $0.25/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder:free 262K context Free input Free output

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...

qwen/qwen3-coder-plus 1M context $0.65/M input $3.25/M output

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b:batch 1.01M context $2/M input $6/M output

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

qwen/qwen3.6-flash 1M context $0.188/M input $1.125/M output

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

qwen/qwen3.7-plus 1M context $0.32/M input $1.28/M output

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

qwen/qwen3.6-plus 1M context $0.325/M input $1.95/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 1M context $0.42/M input $3/M output

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-14b 131.072K context $0.227/M input $0.91/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder 262.144K context $0.3/M input $1/M output

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output