Alibaba Qwen logo

Alibaba Qwen: Qwen: Qwen3 VL 32B Instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Source-linked Weight access not listed
API record Report
InputT
OutputT
Input price$0.104/M
Output price$0.416/M
Context131.072K
Max output32.768K
Providers0
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
No inference providers are listed

The model record exists, but no source-linked provider offer is available yet.

Submit a source
Specification

Capabilities

Recorded from the source catalog and provider listings.

? Reasoning Unknown
? Tool calling Unknown
? Structured output Unknown
? Attachments Unknown
Vision input Yes
? Open weights Unknown
Model family
Not documented
Knowledge cutoff
Not documented
License
Not documented
Release date
Not documented
Model ID
qwen/qwen3-vl-32b-instruct
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
No published benchmark results

ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.

Read the methodology
Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion0.2704 → 0.41600000000000004
Price Prompt0.0676 → 0.10400000000000001
Price Completion0.41600000000000004 → 0.2704
Price Prompt0.10400000000000001 → 0.0676
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Last verified
Sep 11, 2026
Status
Source-linked
Confidence
Medium
Catalog source
OpenRouter
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/qwen/qwen3-vl-32b-instruct
curl
curl "https://model.kyssta.lol/api/v1/models/qwen/qwen3-vl-32b-instruct"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Qwen: Qwen3 VL 32B Instruct?

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text. It is published by Alibaba Qwen and catalogued here from OpenRouter.

What is the context length of Qwen: Qwen3 VL 32B Instruct?

Qwen: Qwen3 VL 32B Instruct accepts up to 131.072K tokens of context and returns up to 32.768K output tokens.

Does Qwen: Qwen3 VL 32B Instruct support tool calling and structured output?

Provider catalogs list support for image input.

More models from Alibaba Qwen

View all →