Alibaba Qwen logo

Alibaba Qwen: Qwen3 32B

Dense open Qwen model for self-hosted chat, reasoning, and coding

Source-linked qwen Open weights Released 2025-04
API record Report
InputT
OutputT
Input price$0.7/M
Output price$2.8/M
Context131.072K
Max output16.384K
Providers14
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Qwen3 32B
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
OpenRouter qwen/qwen3-32b 131.072K 16.384K $0.08 $0.28 ReasoningToolsJSON Docs ↗
Deep Infra Qwen/Qwen3-32B 40.96K 16.384K $0.08 $0.28 ReasoningToolsJSON Docs ↗
Kilo Gateway qwen/qwen3-32b 40.96K 16.384K $0.08 $0.28 ReasoningToolsJSON Docs ↗
Cortecs qwen3-32b 32K 32K $0.089 $0.312 ReasoningToolsJSON Docs ↗
Abacus Qwen/Qwen3-32B 131.072K 8.192K $0.09 $0.29 ReasoningTools Docs ↗
Jiekou.AI qwen/qwen3-32b-fp8 40.96K 20K $0.1 $0.45 Reasoning Docs ↗
Merge Gateway qwen/qwen3-32b 131.072K 32.768K $0.15 $0.6 ReasoningTools Docs ↗
DigitalOcean alibaba-qwen3-32b 32.768K 32.768K $0.25 $0.55 ReasoningToolsJSON Docs ↗
Alibaba (China) qwen3-32b 131.072K 16.384K $0.287 $1.147 ReasoningTools Docs ↗
Helicone qwen3-32b 131.072K 40.96K $0.29 $0.59 ReasoningTools Docs ↗
Hugging Face Qwen/Qwen3-32B 131.072K 16.384K $0.29 $0.59 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) qwen3-32b 40.96K 16.384K $0.36 $0.87 ReasoningTools Docs ↗
Alibaba qwen3-32b 131.072K 16.384K $0.7 $2.8 ReasoningTools Docs ↗
Pioneer Qwen/Qwen3-32B 131.072K 8.192K $0.9 $0.9 $0.9 ReasoningTools Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.08/M

Across 14 priced providers

Median input $0.2/M

Midpoint of listed rates

Highest input $0.9/M

11.2× the lowest listed rate

Output range $0.28 – $2.8

Per million output tokens

OpenRouter $0.08/M
Deep Infra $0.08/M
Kilo Gateway $0.08/M
Cortecs $0.089/M
Abacus $0.09/M
Jiekou.AI $0.1/M
Merge Gateway $0.15/M
DigitalOcean $0.25/M
Alibaba (China) $0.287/M
Helicone $0.29/M
Hugging Face $0.29/M
DevPass (LLM Gateway) $0.36/M
Alibaba $0.7/M
Pioneer $0.9/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

3 providers list the identical $0.08 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.

Context limits also differ by provider, from 32K to 131.072K tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
× Attachments No
× Vision input No
Open weights Yes
Model family
qwen
Knowledge cutoff
2025-04
License
Not documented
Release date
2025-04
Model ID
alibaba/qwen3-32b

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Aider Polyglot percent correct · 2025-05-08 Source ↗
Aider Polyglot #20 of 31

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against Aider Polyglot, the benchmark with the widest published coverage that this model appears in.

0 47 93 $0.1 $1 $10 Qwen Max — $1.6/M, score 21.8 Qwen3 235B-A22B — $0.7/M, score 59.6 Qwen3 32B — $0.7/M, score 40.0 Claude Haiku 3.5 — $0.8/M, score 28.0 Claude Sonnet 3.7 — $3/M, score 64.9 Claude Opus 4 — $15/M, score 72.0 Claude Sonnet 4 (latest) — $2.898/M, score 61.3 Claude Sonnet 4 — $3/M, score 61.3 Command A — $2.5/M, score 12.0 DeepSeek Chat — $0.147/M, score 70.2 DeepSeek-R1 — $0.7/M, score 56.9 DeepSeek Reasoner — $0.147/M, score 74.2 Gemini 2.5 Flash — $0.3/M, score 55.1 Gemini 2.5 Pro — $1.25/M, score 83.1 Llama 4 Maverick 17B Instruct — $0.14/M, score 15.6 Codestral (latest) — $0.3/M, score 11.1 GPT-4.1 — $2/M, score 52.4 GPT-4.1 mini — $0.4/M, score 32.4 GPT-4.1 nano — $0.1/M, score 8.9 GPT-4o — $2.5/M, score 23.1 GPT-4o (2024-08-06) — $2.5/M, score 23.1 GPT-4o (2024-11-20) — $2.5/M, score 18.2 GPT-4o mini — $0.15/M, score 3.6 GPT-5 — $1.25/M, score 88.0 o1 — $15/M, score 61.7 o3 — $2/M, score 81.3 o3-mini — $1.1/M, score 60.4 o3-pro — $20/M, score 84.9 o4-mini — $1.1/M, score 72.0 Qwen3 32B Input price per million tokens (log scale) Percent correct

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion0.28 → 2.8
Price Prompt0.08 → 0.7
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/alibaba/qwen3-32b
curl
curl "https://model.kyssta.lol/api/v1/models/alibaba/qwen3-32b"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Qwen3 32B?

Dense open Qwen model for self-hosted chat, reasoning, and coding. It is published by Alibaba Qwen and catalogued here from Models.dev.

How much does Qwen3 32B cost?

Listed input pricing starts at $0.08 per million tokens from OpenRouter, rising to $0.9 across 14 listed providers.

What is the context length of Qwen3 32B?

Qwen3 32B accepts up to 131.072K tokens of context and returns up to 16.384K output tokens.

Does Qwen3 32B support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, and reasoning.

Which providers serve Qwen3 32B?

14 providers list this model: OpenRouter, Deep Infra, Kilo Gateway, Cortecs, Abacus, Jiekou.AI and 8 more.

Are the weights for Qwen3 32B open?

Yes. The weights are published and downloadable from Hugging Face.

When was Qwen3 32B released?

The catalog records a release date of 2025-04, last verified Sep 11, 2026.

More models from Alibaba Qwen

View all →