No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Sol family.
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
This model always redirects to the latest model in the GPT Astra family.
No provider description is available for this model yet.
The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Luna family.
Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-west-1/xai.grok-4.6bedrock_mantle/us-gov-west-1/xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| accounts/fireworks/routers/kimi-k3-usfireworks_ai/accounts/fireworks/routers/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — | |||
| us-gov-west-1/google.gemma-4-e2bbedrock_mantle/us-gov-west-1/google.gemma-4-e2b | 128K | $0.048 | $0.096 | — | |||
| us-gov-west-1/google.gemma-4-26b-a4bbedrock_mantle/us-gov-west-1/google.gemma-4-26b-a4b | 256K | $0.156 | $0.48 | — | |||
| us-gov-west-1/google.gemma-4-31bbedrock_mantle/us-gov-west-1/google.gemma-4-31b | 256K | $0.168 | $0.48 | — | |||
| us-gov-west-1/openai.gpt-oss-20bbedrock_mantle/us-gov-west-1/openai.gpt-oss-20b | 131.072K | $0.084 | $0.36 | — | |||
| us-gov-west-1/openai.gpt-oss-120bbedrock_mantle/us-gov-west-1/openai.gpt-oss-120b | 131.072K | $0.18 | $0.72 | — | |||
| us-gov-east-1/xai.grok-4.6bedrock_mantle/us-gov-east-1/xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| us-gov-east-1/openai.gpt-oss-20bbedrock_mantle/us-gov-east-1/openai.gpt-oss-20b | 131.072K | $0.084 | $0.36 | — | |||
| us-gov-east-1/openai.gpt-oss-120bbedrock_mantle/us-gov-east-1/openai.gpt-oss-120b | 131.072K | $0.18 | $0.72 | — | |||
| Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct | 32.768K | $0.36 | $0.4 | — | |||
| accounts/fireworks/models/muse-glimmer-30bfireworks_ai/accounts/fireworks/models/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| us-east-1/moonshotai.kimi-k2.5bedrock/us-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| nemotron-3-ultra-nvfp4fireworks_ai/nemotron-3-ultra-nvfp4 | 262.144K | $0.6 | $2.4 | — | |||
| us-east-1/qwen.qwen3-coder-nextbedrock/us-east-1/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| grok-4.20-multi-agent-0309xai/grok-4.20-multi-agent-0309 | 1M | $1.25 | $2.5 | — | |||
| claude-mythos-previewanthropic/claude-mythos-preview | 1M | $10 | $50 | — | |||
| Mistral: Mixtral 8x22B Instructmistralai/mixtral-8x22b-instruct | 65.536K | $2 | $6 | — | |||
| labs-leanstral-1-5mistral/labs-leanstral-1-5 | 262.144K | — | — | — | |||
| gpt-4ogithub_copilot/gpt-4o | 64K | — | — | — | |||
| us-east-2/deepseek.v3.2bedrock/us-east-2/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| gpt-4o-2024-05-13github_copilot/gpt-4o-2024-05-13 | 64K | — | — | — | |||
| gpt-4o-2024-08-06github_copilot/gpt-4o-2024-08-06 | 64K | — | — | — | |||
| OpenAI: GPT Sol Latest~openai/gpt-sol-latest | 1.05M | $2 | $10 | — | |||
| Z.ai: GLM 5.3z-ai/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| OpenAI: GPT Astra Latest~openai/gpt-astra-latest | 1.05M | $10 | $50 | — | |||
| llama3.2-3bsnowflake/llama3.2-3b | 128K | — | — | — | |||
| OpenAI: o1 (batch)openai/o1:batch | 200K | $7.5 | $30 | — | |||
| llama3.3-70bsnowflake/llama3.3-70b | 128K | $0.72 | $0.72 | — | |||
| mistral-7bsnowflake/mistral-7b | 32K | — | — | — | |||
| mistral-largesnowflake/mistral-large | 32K | — | — | — | |||
| gpt-4o-2024-11-20github_copilot/gpt-4o-2024-11-20 | 64K | — | — | — | |||
| gpt-4o-minigithub_copilot/gpt-4o-mini | 64K | — | — | — | |||
| us-east-2/minimax.minimax-m2.1bedrock/us-east-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| gpt-4o-mini-2024-07-18github_copilot/gpt-4o-mini-2024-07-18 | 64K | — | — | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — | |||
| gpt-5-minigithub_copilot/gpt-5-mini | 128K | — | — | — | |||
| us-gov-east-1/nvidia.nemotron-nano-9b-v2bedrock/us-gov-east-1/nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| us-east-2/minimax.minimax-m2.5bedrock/us-east-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| us-east-2/moonshotai.kimi-k2-thinkingbedrock/us-east-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| FW-Nemotron-3-Ultra-NVFP4azure_ai/fw-nemotron-3-ultra-nvfp4 | 262.144K | $0.6 | $2.4 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| Qwen: Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct | 32.768K | $0.1 | $0.2 | — | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 262.144K | $0.17 | $0.25 | — | |||
| ai21.jamba-1-5-large-v1:0bedrock/ai21.jamba-1-5-large-v1:0 | 256K | $2 | $8 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| FW-MiniMax-M3azure_ai/fw-minimax-m3 | 512K | $0.33 | $1.32 | — |