Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Luna family.
No provider description is available for this model yet.
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| us/gpt-4o-mini-2024-07-18azure/us/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| us/gpt-5-2025-08-07azure/us/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| @cf/meta/llama-3.3-70b-instruct-fp8-fastcloudflare/@cf/meta/llama-3.3-70b-instruct-fp8-fast | 24K | $0.293 | $2.253 | — | |||
| us/gpt-5-mini-2025-08-07azure/us/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| us/gpt-5-nano-2025-08-07azure/us/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| ai21.j2-mid-v1bedrock/ai21.j2-mid-v1 | 8.191K | $12.5 | $12.5 | — | |||
| claude-sonnet-4-6azure_ai/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| @cf/ibm-granite/granite-4.0-h-microcloudflare/@cf/ibm-granite/granite-4.0-h-micro | 131K | $0.017 | $0.112 | — | |||
| us/gpt-5.1-chatazure/us/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| us/o1-mini-2024-09-12azure/us/o1-mini-2024-09-12 | 128K | $1.21 | $4.84 | — | |||
| @cf/qwen/qwen2.5-coder-32b-instructcloudflare/@cf/qwen/qwen2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 262.144K | $0.15 | $0.15 | — | |||
| daybreak-blue-latestopenai/daybreak-blue-latest | 1.05M | $5 | $30 | — | |||
| us/o1-preview-2024-09-12azure/us/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| qwen-3-32bcerebras/qwen-3-32b | 128K | $0.4 | $0.8 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| @cf/deepseek-ai/deepseek-r1-distill-qwen-32bcloudflare/@cf/deepseek-ai/deepseek-r1-distill-qwen-32b | 80K | $0.497 | $4.881 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/o4-mini-2025-04-16azure/us/o4-mini-2025-04-16 | 200K | $1.21 | $4.84 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| claude-sonnet-4-5azure_ai/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| Sakana: Namazusakana/namazu | 262.144K | $0.95 | $4 | — | |||
| daybreak-red-latestopenai/daybreak-red-latest | 400K | $12.5 | $75 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| @cf/moonshotai/kimi-k2.7-codecloudflare/@cf/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| openai/gpt-chat-latestopenrouter/openai/gpt-chat-latest | 400K | $5 | $30 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| claude-opus-4-1azure_ai/claude-opus-4-1 | 200K | $15 | $75 | — | |||
| Inflection: Inflection 3 Productivityinflection/inflection-3-productivity | 8K | $2.5 | $10 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| claude-opus-4-8azure_ai/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| qwen/qwen3.8-max-0902openrouter/qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| us-east-2/moonshotai.kimi-k2-thinkingbedrock/us-east-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| us-east-2/minimax.minimax-m2.5bedrock/us-east-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| us-east-1/1-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| Baidu: ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b | 123K | $0.42 | $1.25 | — | |||
| openai/gpt-6-astra-proopenrouter/openai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| claude-opus-4-7azure_ai/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| gpt-5-minigithub_copilot/gpt-5-mini | 128K | — | — | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| openai/gpt-5.6-terra-proopenrouter/openai/gpt-5.6-terra-pro | 1.05M | $2 | $12 | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — |