No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
No provider description is available for this model yet.
OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| eu-west-3/mistral.mixtral-8x7b-instruct-v0:1bedrock/eu-west-3/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.59 | $0.91 | — | |||
| eu-south-1/minimax.minimax-m2.1bedrock/eu-south-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-south-1/minimax.minimax-m2.5bedrock/eu-south-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| eu-south-1/qwen.qwen3-coder-nextbedrock/eu-south-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| llama3.2-1bsnowflake/llama3.2-1b | 128K | — | — | — | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 1.04758M | $0.2 | $0.8 | — | |||
| zai-glm-4.6cerebras/zai-glm-4.6 | 128K | $2.25 | $2.75 | — | |||
| OpenAI: o3 Mini High (batch)openai/o3-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| gemini-robotics-er-2-previewgemini/gemini-robotics-er-2-preview | 131.072K | $2 | $10 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| gemini-robotics-er-1.6-previewgemini/gemini-robotics-er-1.6-preview | 131.072K | $1 | $5 | — | |||
| nvidia/Ising-Calibration-1-35B-A3Bnvidia/Ising-Calibration-1-35B-A3B | Not documented | — | — | — | |||
| invoke/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/invoke/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| Qwen: Qwen3 Next 80B A3B Instruct (free)qwen/qwen3-next-80b-a3b-instruct:free | 262.144K | Free | Free | — | |||
| sa-east-1/meta.llama3-70b-instruct-v1:0bedrock/sa-east-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $4.45 | $5.88 | — | |||
| sa-east-1/meta.llama3-8b-instruct-v1:0bedrock/sa-east-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.5 | $1.01 | — | |||
| llama2:70bollama/llama2:70b | 4.096K | — | — | — | |||
| sa-east-1/deepseek.v3.2bedrock/sa-east-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| llama2:7bollama/llama2:7b | 4.096K | — | — | — | |||
| llama3ollama/llama3 | 8.192K | — | — | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| sa-east-1/minimax.minimax-m2.1bedrock/sa-east-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Qwen: Qwen Plus 0728 (thinking)qwen/qwen-plus-2025-07-28:thinking | 1M | $0.26 | $0.78 | — | |||
| gpt-6-astraazure_ai/gpt-6-astra | 922K | $10 | $50 | — | |||
| sa-east-1/minimax.minimax-m2.5bedrock/sa-east-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| sa-east-1/moonshotai.kimi-k2-thinkingbedrock/sa-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| gpt-5.6-cyberopenai/gpt-5.6-cyber | 400K | $12.5 | $75 | — | |||
| sa-east-1/moonshotai.kimi-k2.5bedrock/sa-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| sa-east-1/qwen.qwen3-coder-nextbedrock/sa-east-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| OpenAI: gpt-oss-120b (free)openai/gpt-oss-120b:free | 131.072K | Free | Free | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731deepseek-ai/DeepSeek-V4-Flash-0731 | Not documented | — | — | — | |||
| us-east-1/1-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| daybreak-red-latestopenai/daybreak-red-latest | 400K | $12.5 | $75 | — | |||
| daybreak-blue-latestopenai/daybreak-blue-latest | 1.05M | $5 | $30 | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| us-east-1/6-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| chat-latestopenai/chat-latest | 400K | $5 | $30 | — | |||
| zai-glm-5-2mistral/zai-glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| deepseek/deepseek-v4-proopenrouter/deepseek/deepseek-v4-pro | 1.04858M | $1.32 | $3.96 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| deepseek/deepseek-v4-pro-0813openrouter/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| llama3.1ollama/llama3.1 | 8.192K | — | — | — | |||
| us-gov-west-1/amazon.nova-lite-v1:0bedrock/us-gov-west-1/amazon.nova-lite-v1:0 | 300K | $0.072 | $0.288 | — | |||
| us-gov-west-1/amazon.nova-micro-v1:0bedrock/us-gov-west-1/amazon.nova-micro-v1:0 | 128K | $0.042 | $0.168 | — | |||
| us-east-1/anthropic.claude-instant-v1bedrock/us-east-1/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — |