No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Sol family.
No provider description is available for this model yet.
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
This model always redirects to the latest model in the GPT Astra family.
No provider description is available for this model yet.
This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.
The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
This model always redirects to the latest model in the GPT Luna family.
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-east-1/openai.gpt-oss-120bbedrock_mantle/us-gov-east-1/openai.gpt-oss-120b | 131.072K | $0.18 | $0.72 | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| us-east-1/moonshotai.kimi-k2-thinkingbedrock/us-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| us-east-1/moonshotai.kimi-k2.5bedrock/us-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| us-east-1/qwen.qwen3-coder-nextbedrock/us-east-1/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| grok-4.20-multi-agent-0309xai/grok-4.20-multi-agent-0309 | 1M | $1.25 | $2.5 | — | |||
| claude-mythos-previewanthropic/claude-mythos-preview | 1M | $10 | $50 | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| gemini-robotics-er-2-streaming-previewgemini/gemini-robotics-er-2-streaming-preview | Not documented | $2 | $10 | — | |||
| meta-llama/Llama-4-Maverick-17B-128E-Instructmeta-llama/Llama-4-Maverick-17B-128E-Instruct | Not documented | — | — | — | |||
| labs-leanstral-1-5mistral/labs-leanstral-1-5 | 262.144K | — | — | — | |||
| meta-llama/Llama-4-Scout-17B-16E-Instructmeta-llama/Llama-4-Scout-17B-16E-Instruct | Not documented | — | — | — | |||
| meta-llama/Llama-Guard-4-12Bmeta-llama/Llama-Guard-4-12B | Not documented | — | — | — | |||
| meta-llama/Llama-4-Maverick-17B-128Emeta-llama/Llama-4-Maverick-17B-128E | Not documented | — | — | — | |||
| meta-llama/Llama-4-Scout-17B-16Emeta-llama/Llama-4-Scout-17B-16E | Not documented | — | — | — | |||
| gpt-41-copilotgithub_copilot/gpt-41-copilot | Not documented | — | — | — | |||
| gpt-4ogithub_copilot/gpt-4o | 64K | — | — | — | |||
| us-east-2/deepseek.v3.2bedrock/us-east-2/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| gpt-4o-2024-05-13github_copilot/gpt-4o-2024-05-13 | 64K | — | — | — | |||
| gpt-4o-2024-08-06github_copilot/gpt-4o-2024-08-06 | 64K | — | — | — | |||
| OpenAI: GPT Sol Latest~openai/gpt-sol-latest | 1.05M | $2 | $10 | — | |||
| bigcode/deepseekcoder-33b-codeqwen-align-subsetbigcode/deepseekcoder-33b-codeqwen-align-subset | Not documented | — | — | — | |||
| Z.ai: GLM 5.3z-ai/glm-5.3 | 1.04858M | $1.4 | $4.4 | — | |||
| OpenAI: GPT Astra Latest~openai/gpt-astra-latest | 1.05M | $10 | $50 | — | |||
| llama3.2-3bsnowflake/llama3.2-3b | 128K | — | — | — | |||
| OpenAI: GPT-3.5 Turbo Instructopenai/gpt-3.5-turbo-instruct | 4.095K | $1.5 | $2 | — | |||
| OpenAI: o1 (batch)openai/o1:batch | 200K | $7.5 | $30 | — | |||
| llama3.3-70bsnowflake/llama3.3-70b | 128K | $0.72 | $0.72 | — | |||
| mistral-7bsnowflake/mistral-7b | 32K | — | — | — | |||
| mistral-largesnowflake/mistral-large | 32K | — | — | — | |||
| gpt-4o-2024-11-20github_copilot/gpt-4o-2024-11-20 | 64K | — | — | — | |||
| gpt-4o-minigithub_copilot/gpt-4o-mini | 64K | — | — | — | |||
| us-east-2/minimax.minimax-m2.1bedrock/us-east-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| gpt-4o-mini-2024-07-18github_copilot/gpt-4o-mini-2024-07-18 | 64K | — | — | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — | |||
| gpt-5-minigithub_copilot/gpt-5-mini | 128K | — | — | — | |||
| us-east-2/minimax.minimax-m2.5bedrock/us-east-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| us-east-2/moonshotai.kimi-k2-thinkingbedrock/us-east-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-it | 32.768K | $0.06 | $0.12 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 262.144K | $0.15 | $0.15 | — | |||
| Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.9 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| Qwen: Qwen3.5-9B (batch)qwen/qwen3.5-9b:batch | 262.144K | $0.17 | $0.25 | — | |||
| FW-GLM-5.2azure_ai/fw-glm-5.2 | 1.04858M | $1.54 | $4.84 | — | |||
| Amazon: Nova Pro 1.0amazon/nova-pro-v1 | 300K | $0.8 | $3.2 | — | |||
| ai21.j2-ultra-v1bedrock/ai21.j2-ultra-v1 | 8.191K | $18.8 | $18.8 | — | |||
| FW-GLM-5.1azure_ai/fw-glm-5.1 | 202.8K | $1.54 | $4.84 | — |