No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
No provider description is available for this model yet.
No provider description is available for this model yet.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| ByteDance Seed: Seed 2.1 Turbobytedance-seed/seed-2-1-turbo | 262.144K | $0.5 | $2.5 | — | |||
| us-west-2/mistral.mistral-large-2402-v1:0bedrock/us-west-2/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| us-gov-east-1/nvidia.nemotron-super-3-120bbedrock/us-gov-east-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us-west-2/mistral.mixtral-8x7b-instruct-v0:1bedrock/us-west-2/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.45 | $0.7 | — | |||
| gpt-5.4-miniazure_ai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| us-gov-east-1/openai.gpt-oss-20b-1:0bedrock/us-gov-east-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| deepseek/deepseek-v4-pro-0813openrouter/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| gpt-5.4-mini-2026-03-17azure_ai/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| us-gov-east-1/openai.gpt-oss-120b-1:0bedrock/us-gov-east-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| Qwen: Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-5bedrock/us-gov-east-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| gpt-5.4-nano-2026-03-17azure_ai/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| us-gov-east-1/anthropic.claude-opus-4-8bedrock/us-gov-east-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| eu.anthropic.claude-opus-4-8bedrock_converse/eu.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| eu/gpt-4o-2024-08-06azure/eu/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| eu/gpt-4o-2024-11-20azure/eu/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| @cf/meta/llama-3.2-1b-instructcloudflare/@cf/meta/llama-3.2-1b-instruct | 60K | $0.027 | $0.201 | — | |||
| eu/gpt-4o-mini-2024-07-18azure/eu/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| eu/gpt-5-2025-08-07azure/eu/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| eu/gpt-5-mini-2025-08-07azure/eu/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| eu/gpt-5.1azure/eu/gpt-5.1 | 272K | $1.38 | $11 | — | |||
| eu/gpt-5.1-chatazure/eu/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| eu/gpt-5-nano-2025-08-07azure/eu/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| us.anthropic.claude-opus-4-8bedrock_converse/us.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| eu/o1-2024-12-17azure/eu/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| us-gov-west-1/xai.grok-4.3bedrock_mantle/us-gov-west-1/xai.grok-4.3 | 131.072K | $1.5 | $3 | — | |||
| eu/o1-preview-2024-09-12azure/eu/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| Qwen: Qwen3.8 27Bqwen/qwen3.8-27b | 262.144K | $0.214 | $2.55 | — | |||
| us-gov/gpt-5.1azure/us-gov/gpt-5.1 | 272K | $1.719 | $13.75 | — | |||
| us-gov/o3-miniazure/us-gov/o3-mini | 200K | $1.513 | $6.05 | — | |||
| Z.ai: GLM 5.2 (free)z-ai/glm-5.2:free | 256K | Free | Free | — | |||
| deepseek/deepseek-v4-proopenrouter/deepseek/deepseek-v4-pro | 1.04858M | $1.32 | $3.96 | — | |||
| global-standard/gpt-4o-2024-08-06azure/global-standard/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| global-standard/gpt-4o-2024-11-20azure/global-standard/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| Qwen: Qwen3 VL 8B Thinkingqwen/qwen3-vl-8b-thinking | 131.072K | $0.18 | $2.1 | — | |||
| global-standard/gpt-4o-miniazure/global-standard/gpt-4o-mini | 128K | $0.15 | $0.6 | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| global/gpt-5.1azure/global/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| Tencent: Hunyuan A13B Instructtencent/hunyuan-a13b-instruct | 131.072K | $0.14 | $0.57 | — | |||
| global/gpt-5.1-chatazure/global/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| databricks-claude-fable-5databricks/databricks-claude-fable-5 | 1M | $10 | $50 | — | |||
| Meta: Llama 3.3 70B Instruct (free)meta-llama/llama-3.3-70b-instruct:free | 65.536K | Free | Free | — | |||
| gpt-3.5-turbo-0125azure/gpt-3.5-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| global.anthropic.claude-opus-4-8bedrock_converse/global.anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| gpt-35-turbo-0125azure/gpt-35-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — |