142 models
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-397B-A17B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-122B-A10B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-27B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.6-35B-A3B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.6-27B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-35B-A3B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-9B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-35B-A3B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-0.8B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-2B-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-4B-Base Not documented context Input not listed Output not listed

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b:batch 1.01M context $2/M input $6/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-0.8B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-2B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-4B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3.5-9B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3-Coder-Next Not documented context Input not listed Output not listed

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

qwen/qwen3.6-flash 1M context $0.188/M input $1.125/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen3-VL-8B-Thinking Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3Guard-Gen-8B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3Guard-Gen-0.6B Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

Qwen/Qwen3Guard-Gen-4B Not documented context Input not listed Output not listed

Qwen3.8 Max (0803) is the August 3, 2026 checkpoint of Qwen3.8 Max, the flagship model in Alibaba's Qwen3.8 series and the general-availability successor to the Qwen3.8 Max Preview. It is...

qwen/qwen3.8-max 1M context $2/M input $6/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 1M context $0.42/M input $3/M output

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

qwen/qwen-plus 1M context $0.26/M input $0.78/M output

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

qwen/qwen2.5-vl-72b-instruct 128K context $0.8/M input $1/M output

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...

qwen/qwen3-vl-235b-a22b-instruct 131.072K context $0.21/M input $1.9/M output

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

qwen/qwen3-next-80b-a3b-thinking 262.144K context $0.15/M input $1.2/M output

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....

qwen/qwen3-vl-235b-a22b-thinking 131.072K context $0.4/M input $4/M output

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b 1M context $2/M input $6/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen-Drive-1.0-4B Not documented context Input not listed Output not listed

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

qwen/qwen3-vl-32b-instruct 131.072K context $0.104/M input $0.416/M output

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28 1M context $0.26/M input $0.78/M output
Open weights

No provider description is available for this model yet.

Qwen/Qwen-Image-2512 Not documented context Input not listed Output not listed

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

qwen/qwen3-30b-a3b-thinking-2507 81.92K context $0.2/M input $2.4/M output

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

qwen/qwen3.7-flash 1M context $0.03/M input $0.13/M output

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-72b-instruct 32.768K context $0.36/M input $0.4/M output

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

qwen/qwen3-vl-8b-instruct 131.072K context $0.117/M input $0.455/M output

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

qwen/qwen3.7-max 1M context $1.475/M input $4.425/M output

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

qwen/qwen3.6-35b-a3b 262.144K context $0.1/M input $0.9/M output