No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window. Devstral 2 supports exploring...
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
This model always redirects to the latest model in the GPT Luna family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| command-r-pluscohere_chat/command-r-plus | 128K | $2.5 | $10 | — | |||
| meta-llama/Llama-3.2-1B-Instructmeta-llama/Llama-3.2-1B-Instruct | Not documented | — | — | — | |||
| meta/llama-2-70b-chatreplicate/meta/llama-2-70b-chat | 4.096K | $0.65 | $2.75 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-Vision-Expdeepseek-ai/DeepSeek-V4-Flash-Vision-Exp | Not documented | — | — | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| Mistral: Devstral 2 2512mistralai/devstral-2512 | 262.144K | $0.4 | $2 | — | |||
| meta-llama/Llama-3.2-3Bmeta-llama/Llama-3.2-3B | Not documented | — | — | — | |||
| Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28 | 1M | $0.26 | $0.78 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 131.072K | $0.4 | $2 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| microsoft/Fara1.5-4Bmicrosoft/Fara1.5-4B | Not documented | — | — | — | |||
| google/gemma-4-31B-it-Ultradeepinfra/google/gemma-4-31b-it-ultra | 131.072K | $0.27 | $0.76 | — | |||
| meta-llama/Llama-3.1-8Bmeta-llama/Llama-3.1-8B | Not documented | — | — | — | |||
| meta-llama/Llama-Guard-3-8Bmeta-llama/Llama-Guard-3-8B | Not documented | — | — | — | |||
| meta-llama/Meta-Llama-3-70Bmeta-llama/Meta-Llama-3-70B | Not documented | — | — | — | |||
| microsoft/MagenticBrainmicrosoft/MagenticBrain | Not documented | — | — | — | |||
| meta-llama/Meta-Llama-3-8Bmeta-llama/Meta-Llama-3-8B | Not documented | — | — | — | |||
| OpenAI: GPT-5 Codex (batch)openai/gpt-5-codex:batch | 400K | $0.625 | $5 | — | |||
| ai21.j2-mid-v1bedrock/ai21.j2-mid-v1 | 8.191K | $12.5 | $12.5 | — | |||
| meta-llama/Llama-3.2-90B-Visionmeta-llama/Llama-3.2-90B-Vision | Not documented | — | — | — | |||
| meta-llama/Llama-3.2-11B-Visionmeta-llama/Llama-3.2-11B-Vision | Not documented | — | — | — | |||
| meta-llama/Llama-Guard-3-1Bmeta-llama/Llama-Guard-3-1B | Not documented | — | — | — | |||
| ai21.j2-ultra-v1bedrock/ai21.j2-ultra-v1 | 8.191K | $18.8 | $18.8 | — | |||
| ai21.jamba-1-5-large-v1:0bedrock/ai21.jamba-1-5-large-v1:0 | 256K | $2 | $8 | — | |||
| llama-2-70b-chatperplexity/llama-2-70b-chat | 4.096K | $0.7 | $2.8 | — | |||
| ai21.jamba-1-5-mini-v1:0bedrock/ai21.jamba-1-5-mini-v1:0 | 256K | $0.2 | $0.4 | — | |||
| microsoft/Mage-VLmicrosoft/Mage-VL | Not documented | — | — | — | |||
| ai21.jamba-instruct-v1:0bedrock/ai21.jamba-instruct-v1:0 | 70K | $0.5 | $0.7 | — | |||
| meta-llama/Llama-Guard-3-1B-INT4meta-llama/Llama-Guard-3-1B-INT4 | Not documented | — | — | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| meta-llama/Llama-3.1-405B-Instructmeta-llama/Llama-3.1-405B-Instruct | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-405Bmeta-llama/Llama-3.1-405B | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-70Bmeta-llama/Llama-3.1-70B | Not documented | — | — | — | |||
| us.writer.palmyra-x4-v1:0bedrock_converse/us.writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| us.writer.palmyra-x5-v1:0bedrock_converse/us.writer.palmyra-x5-v1:0 | 1M | $0.6 | $6 | — | |||
| Qwen/Qwen3.8-27Bdeepinfra/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| microsoft/Mage-Flow-Turbomicrosoft/Mage-Flow-Turbo | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-8B-Instructmeta-llama/Llama-3.1-8B-Instruct | Not documented | — | — | — | |||
| nvidia/Qwen-Image-Flashnvidia/Qwen-Image-Flash | Not documented | — | — | — | |||
| writer.palmyra-x4-v1:0bedrock_converse/writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| writer.palmyra-x5-v1:0bedrock_converse/writer.palmyra-x5-v1:0 | 1M | $0.6 | $6 | — | |||
| amazon.nova-lite-v1:0bedrock_converse/amazon.nova-lite-v1:0 | 300K | $0.06 | $0.24 | — | |||
| amazon.nova-2-lite-v1:0bedrock_converse/amazon.nova-2-lite-v1:0 | 1M | $0.3 | $2.5 | — | |||
| MiniMaxAI/MiniMax-M2.7deepinfra/minimaxai/minimax-m2.7 | 196.608K | $0.25 | $1 | — | |||
| meta-llama/Llama-Guard-3-8B-INT8meta-llama/Llama-Guard-3-8B-INT8 | Not documented | — | — | — | |||
| meta-llama/Meta-Llama-Guard-2-8Bmeta-llama/Meta-Llama-Guard-2-8B | Not documented | — | — | — | |||
| nvidia/Riva-Translate-4B-Instructnvidia/Riva-Translate-4B-Instruct | Not documented | — | — | — | |||
| amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| codellama-70b-instructperplexity/codellama-70b-instruct | 16.384K | $0.7 | $2.8 | — |