No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
No provider description is available for this model yet.
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| claude-sonnet-4-5azure_ai/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instruct-FP8gmi/qwen/qwen3-vl-235b-a22b-instruct-fp8 | 262.144K | $0.3 | $1.4 | — | |||
| @cf/deepseek-ai/deepseek-r1-distill-qwen-32bcloudflare/@cf/deepseek-ai/deepseek-r1-distill-qwen-32b | 80K | $0.497 | $4.881 | — | |||
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| claude-sonnet-4-6azure_ai/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| computer-use-previewazure/computer-use-preview | 8.192K | $3 | $12 | — | |||
| containerazure/container | Not documented | — | — | — | |||
| gemini-3.8-flashvertex_ai/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| gpt-oss-120bazure_ai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| OpenAI: GPT-3.5 Turbo (batch)openai/gpt-3.5-turbo:batch | 16.385K | $0.25 | $0.75 | — | |||
| @cf/meta/llama-3.1-8b-instruct-fp8cloudflare/@cf/meta/llama-3.1-8b-instruct-fp8 | 32K | $0.152 | $0.287 | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 262.144K | $0.95 | $4 | — | |||
| gpt-5.5azure_ai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| gemini-3.8-flashgemini/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| DeepSeek: DeepSeek V4 Flash Vision Exp (batch)deepseek/deepseek-v4-flash-vision-exp:batch | 1.04858M | $0.11 | $0.33 | — | |||
| us-gov-west-1/meta.llama3-70b-instruct-v1:0bedrock/us-gov-west-1/meta.llama3-70b-instruct-v1:0 | 8K | $2.65 | $3.5 | — | |||
| us-gov-west-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-west-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| gemini-3.8-flashvertex_ai-language-models/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-east-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| deepseek-ai/ESFT-token-intent-litedeepseek-ai/ESFT-token-intent-lite | Not documented | — | — | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| us-west-1/meta.llama3-70b-instruct-v1:0bedrock/us-west-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.65 | $3.5 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 400K | $0.625 | $5 | — | |||
| us-gov.anthropic.claude-opus-4-8bedrock_converse/us-gov.anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| us-west-1/meta.llama3-8b-instruct-v1:0bedrock/us-west-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.3 | $0.6 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| accounts/fireworks/models/glm-5p3-flashfireworks_ai/accounts/fireworks/models/glm-5p3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| us-west-2/1-month-commitment/anthropic.claude-v2:1bedrock/us-west-2/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| deepseek-v4-flash-0731scaleway/deepseek-v4-flash-0731 | 256K | $0.4 | $0.8 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| kimi-k2.7-codeazure_ai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v2:1bedrock/us-west-2/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-west-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| us-gov-west-1/nvidia.nemotron-super-3-120bbedrock/us-gov-west-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us-west-2/anthropic.claude-instant-v1bedrock/us-west-2/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| us-gov-west-1/openai.gpt-oss-20b-1:0bedrock/us-gov-west-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| us-gov-west-1/openai.gpt-oss-120b-1:0bedrock/us-gov-west-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| Qwen: Qwen3 VL 30B A3B Thinkingqwen/qwen3-vl-30b-a3b-thinking | 131.072K | $0.2 | $2.4 | — |