No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Luna family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
No provider description is available for this model yet.
No provider description is available for this model yet.
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| ai21.jamba-1-5-mini-v1:0bedrock/ai21.jamba-1-5-mini-v1:0 | 256K | $0.2 | $0.4 | — | |||
| us-gov-east-1/anthropic.claude-fable-5-1bedrock/us-gov-east-1/anthropic.claude-fable-5-1 | 1M | $12 | $60 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-9b-v2bedrock/us-gov-west-1/nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| eu/gpt-4o-2024-11-20azure/eu/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| eu/gpt-4o-2024-08-06azure/eu/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| GLM-5.2scx-ai/glm-5.2 | 1.04858M | $0.61 | $1.98 | — | |||
| us-gov-east-1/anthropic.claude-opus-4-8bedrock/us-gov-east-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| us-gov-west-1/amazon.nova-pro-v1:0bedrock/us-gov-west-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| ai21.jamba-1-5-large-v1:0bedrock/ai21.jamba-1-5-large-v1:0 | 256K | $2 | $8 | — | |||
| gpt-5.4-nano-2026-03-17azure_ai/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-5bedrock/us-gov-east-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| gpt-audio-miniazure/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| gpt-5.4-nanoazure_ai/gpt-5.4-nano | 272K | $0.2 | $1.25 | — | |||
| us-gov-east-1/openai.gpt-oss-120b-1:0bedrock/us-gov-east-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| au.anthropic.claude-opus-4-7bedrock_converse/au.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| accounts/fireworks/routers/glm-5p2-fast-usfireworks_ai/accounts/fireworks/routers/glm-5p2-fast-us | 1.04858M | $2.1 | $6.6 | — | |||
| gpt-5.4-mini-2026-03-17azure_ai/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| us-gov-east-1/openai.gpt-oss-20b-1:0bedrock/us-gov-east-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| gpt-5.4-miniazure_ai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| ByteDance Seed: Seed 2.1 Turbobytedance-seed/seed-2-1-turbo | 262.144K | $0.5 | $2.5 | — | |||
| Z.ai: GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch | 1.04858M | $0.075 | $0.25 | — | |||
| us-gov-east-1/nvidia.nemotron-super-3-120bbedrock/us-gov-east-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-east-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| google.gemma-3-12b-itbedrock_converse/google.gemma-3-12b-it | 128K | $0.09 | $0.29 | — | |||
| gpt-5.4-2026-03-05azure_ai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| Meta: Llama 3.2 3B Instruct (free)meta-llama/llama-3.2-3b-instruct:free | 131.072K | Free | Free | — | |||
| zai-org/GLM-4.7-FP8gmi/zai-org/glm-4.7-fp8 | 202.752K | $0.4 | $2 | — | |||
| accounts/fireworks/routers/glm-5p2-fastfireworks_ai/accounts/fireworks/routers/glm-5p2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| muse-glimmer-30bfireworks_ai/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| glm-5p2-fast-usfireworks_ai/glm-5p2-fast-us | 1.04858M | $2.1 | $6.6 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-east-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instruct-FP8gmi/qwen/qwen3-vl-235b-a22b-instruct-fp8 | 262.144K | $0.3 | $1.4 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 262.144K | $0.25 | $0.75 | — | |||
| us-gov-west-1/openai.gpt-oss-120b-1:0bedrock/us-gov-west-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| us-gov-west-1/openai.gpt-oss-20b-1:0bedrock/us-gov-west-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov-west-1/nvidia.nemotron-super-3-120bbedrock/us-gov-west-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-west-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| eu.anthropic.claude-opus-4-7bedrock_converse/eu.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-east-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| kimi-k2.7-codeazure_ai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| deepseek-v4-flash-0731scaleway/deepseek-v4-flash-0731 | 256K | $0.4 | $0.8 | — |