Alibaba Qwen logo

Alibaba Qwen: Qwen3.8 2.4T A95B

Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows

Source-linked qwen Open weights qwen3.8-max Released 2026-08-12
API record Report
InputT
OutputT
Input price$2/M
Output price$6/M
Context262.144K
Max output131.072K
Providers12
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Qwen3.8 2.4T A95B
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Fireworks AI accounts/fireworks/models/qwen3p8-2p4t-a95b 262.144K 131.072K $2 $6 $0.25 ReasoningToolsJSON Docs ↗
Charm Hyper qwen3.8-2.4t-a95b 1M 128K $2 $6 $0.25 ReasoningToolsJSON Docs ↗
Vercel AI Gateway alibaba/qwen3.8-2.4t-a95b 262.144K 128K $2 $6 $0.25 ReasoningToolsJSON Docs ↗
Kilo Gateway qwen/qwen3.8-2.4t-a95b 1M 131.072K $2 $6 $0.25 ReasoningToolsJSON Docs ↗
Requesty qwen3.8-2.4T-A95B 262.144K 262.144K $2 $6 $0.2 ReasoningToolsJSON Docs ↗
Deep Infra Qwen/Qwen3.8-2.4T-A95B 262.144K 131.072K $2 $6 $0.2 ReasoningToolsJSON Docs ↗
OpenRouter qwen/qwen3.8-2.4t-a95b 1.04858M 131.072K $2 $6 $0.25 ReasoningToolsJSON Docs ↗
AIHubMix qwen3.8-2.4t-a95b 262K 262K $2 $6 $0.5 ReasoningToolsJSON Docs ↗
Eden AI qwen/qwen3.8-2.4t-a95b 1M 131.072K $2 $6 $0.25 ReasoningToolsJSON Docs ↗
Cortecs qwen3.8-2.4t-a95b 262.144K 262.144K $2.5 $6 $0.625 ReasoningToolsJSON Docs ↗
Merge Gateway qwen/qwen3.8-2.4t-a95b 262.144K 1.01M $2.5 $6.25 $0.5 Reasoning Docs ↗
Hugging Face Qwen/Qwen3.8-2.4T-A95B 262.144K 131.072K $2.5 $6.25 ReasoningToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $2/M

Across 12 priced providers

Median input $2/M

Midpoint of listed rates

Highest input $2.5/M

1.2× the lowest listed rate

Output range $6 – $6.25

Per million output tokens

Fireworks AI $2/M
Charm Hyper $2/M
Vercel AI Gateway $2/M
Kilo Gateway $2/M
Requesty $2/M
Deep Infra $2/M
OpenRouter $2/M
AIHubMix $2/M
Eden AI $2/M
Cortecs $2.5/M
Merge Gateway $2.5/M
Hugging Face $2.5/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

9 providers list the identical $2 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.

Context limits also differ by provider, from 262K to 1.04858M tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
× Attachments No
× Vision input No
Open weights Yes
Model family
qwen
Knowledge cutoff
Not documented
License
qwen3.8-max
Release date
2026-08-12
Model ID
alibaba/qwen3.8-2.4t-a95b

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
No published benchmark results

ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.

Read the methodology
Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion5.4 → 6.0
Price Prompt1.8 → 2.0
Price Completion6.0 → 5.4
Price Prompt2.0 → 1.8
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/alibaba/qwen3.8-2.4t-a95b
curl
curl "https://model.kyssta.lol/api/v1/models/alibaba/qwen3.8-2.4t-a95b"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Qwen3.8 2.4T A95B?

Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows. It is published by Alibaba Qwen and catalogued here from Models.dev.

How much does Qwen3.8 2.4T A95B cost?

Listed input pricing starts at $2 per million tokens from Fireworks AI, rising to $2.5 across 12 listed providers.

What is the context length of Qwen3.8 2.4T A95B?

Qwen3.8 2.4T A95B accepts up to 262.144K tokens of context and returns up to 131.072K output tokens.

Does Qwen3.8 2.4T A95B support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, and reasoning.

Which providers serve Qwen3.8 2.4T A95B?

12 providers list this model: Fireworks AI, Charm Hyper, Vercel AI Gateway, Kilo Gateway, Requesty, Deep Infra and 6 more.

Are the weights for Qwen3.8 2.4T A95B open?

Yes. The weights are published and downloadable from Hugging Face under the qwen3.8-max license.

When was Qwen3.8 2.4T A95B released?

The catalog records a release date of 2026-08-12, last verified Sep 11, 2026.

More models from Alibaba Qwen

View all →