Alibaba Qwen logo

Alibaba Qwen: Qwen Flash

Efficient Qwen model for fast chat, extraction, and high-volume workloads

Source-linked qwen Closed weights Released 2025-07-28
API record Report
InputT
OutputT
Input price$0.05/M
Output price$0.4/M
Context1M
Max output32.768K
Providers8
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Qwen Flash
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Ofox qwen/qwen-flash 1M 32K $0.022 $0.22 $0.004 ReasoningTools Docs ↗
Ofox bailian/qwen-flash 1M 32.768K $0.022 $0.22 $0.004 ReasoningTools Docs ↗
Alibaba (China) qwen-flash 1M 32.768K $0.022 $0.216 ReasoningTools Docs ↗
Merge Gateway qwen/qwen-flash 1M 250K $0.022 $0.216 $0.004 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) qwen-flash 1M 32.768K $0.05 $0.4 $0.01 ReasoningTools Docs ↗
LLMTR qwen/qwen-flash 1M 32.768K $0.05 $0.4 ReasoningTools Docs ↗
Alibaba qwen-flash 1M 32.768K $0.05 $0.4 ReasoningTools Docs ↗
LLM Gateway alibaba/qwen-flash 1M 32K $0.05 $0.4 $0.01 Tools Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.022/M

Across 8 priced providers

Median input $0.036/M

Midpoint of listed rates

Highest input $0.05/M

2.3× the lowest listed rate

Output range $0.216 – $0.4

Per million output tokens

Ofox $0.022/M
Ofox $0.022/M
Alibaba (China) $0.022/M
Merge Gateway $0.022/M
DevPass (LLM Gateway) $0.05/M
LLMTR $0.05/M
Alibaba $0.05/M
LLM Gateway $0.05/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

4 providers list the identical $0.022 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.

Specification

Capabilities

Recorded from the source catalog and provider listings.

× Reasoning No
Tool calling Yes
× Structured output No
× Attachments No
× Vision input No
× Open weights No
Model family
qwen
Knowledge cutoff
2024-04
License
Not documented
Release date
2025-07-28
Model ID
alibaba/qwen-flash
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
No published benchmark results

ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.

Read the methodology
Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion0.216 → 0.4
Price Prompt0.022 → 0.05
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/alibaba/qwen-flash
curl
curl "https://model.kyssta.lol/api/v1/models/alibaba/qwen-flash"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Qwen Flash?

Efficient Qwen model for fast chat, extraction, and high-volume workloads. It is published by Alibaba Qwen and catalogued here from Models.dev.

How much does Qwen Flash cost?

Listed input pricing starts at $0.022 per million tokens from Ofox, rising to $0.05 across 8 listed providers.

What is the context length of Qwen Flash?

Qwen Flash accepts up to 1M tokens of context and returns up to 32.768K output tokens.

Does Qwen Flash support tool calling and structured output?

Provider catalogs list support for tool calling.

Which providers serve Qwen Flash?

7 providers list this model: Ofox, Alibaba (China), Merge Gateway, DevPass (LLM Gateway), LLMTR, Alibaba and 1 more.

Are the weights for Qwen Flash open?

No. This model is served through hosted APIs only.

When was Qwen Flash released?

The catalog records a release date of 2025-07-28, last verified Sep 11, 2026.

More models from Alibaba Qwen

View all →