No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
No provider description is available for this model yet.
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Luna family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| mistral-largesnowflake/mistral-large | 32K | — | — | — | |||
| gpt-4o-2024-11-20github_copilot/gpt-4o-2024-11-20 | 64K | — | — | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813deepseek-ai/DeepSeek-V4-Pro-0813 | Not documented | — | — | — | |||
| gpt-4o-minigithub_copilot/gpt-4o-mini | 64K | — | — | — | |||
| meta-llama/Llama-3.2-1B-Instruct-QLORA_INT4_EO8meta-llama/Llama-3.2-1B-Instruct-QLORA_INT4_EO8 | Not documented | — | — | — | |||
| gemini-robotics-er-2-previewgemini/gemini-robotics-er-2-preview | 131.072K | $2 | $10 | — | |||
| us-east-2/minimax.minimax-m2.1bedrock/us-east-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| gpt-4o-mini-2024-07-18github_copilot/gpt-4o-mini-2024-07-18 | 64K | — | — | — | |||
| databricks-qwen3-next-80b-a3b-instructdatabricks/databricks-qwen3-next-80b-a3b-instruct | Not documented | $0.15 | $1.2 | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — | |||
| OpenAI: o3 Mini High (batch)openai/o3-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| meta-llama/Llama-3.2-1Bmeta-llama/Llama-3.2-1B | Not documented | — | — | — | |||
| databricks-qwen35-122b-a10bdatabricks/databricks-qwen35-122b-a10b | 262.144K | $0.22 | $2.2 | — | |||
| us-east-2/moonshotai.kimi-k2-thinkingbedrock/us-east-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| Qwen: Qwen3 VL 235B A22B Thinkingqwen/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| meta-llama/Llama-3.2-1B-Instructmeta-llama/Llama-3.2-1B-Instruct | Not documented | — | — | — | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 | $15 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| meta-llama/Llama-3.2-3Bmeta-llama/Llama-3.2-3B | Not documented | — | — | — | |||
| zai-glm-4.6cerebras/zai-glm-4.6 | 128K | $2.25 | $2.75 | — | |||
| microsoft/Fara1.5-4Bmicrosoft/Fara1.5-4B | Not documented | — | — | — | |||
| meta-llama/Llama-3.2-3B-Instructmeta-llama/Llama-3.2-3B-Instruct | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-8Bmeta-llama/Llama-3.1-8B | Not documented | — | — | — | |||
| meta-llama/Llama-Guard-3-8Bmeta-llama/Llama-Guard-3-8B | Not documented | — | — | — | |||
| databricks-grok-4-6databricks/databricks-grok-4-6 | 500K | $2.5 | $7.5 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| microsoft/MagenticBrainmicrosoft/MagenticBrain | Not documented | — | — | — | |||
| meta-llama/Meta-Llama-3-8Bmeta-llama/Meta-Llama-3-8B | Not documented | — | — | — | |||
| ai21.j2-mid-v1bedrock/ai21.j2-mid-v1 | 8.191K | $12.5 | $12.5 | — | |||
| meta-llama/Llama-3.2-90B-Visionmeta-llama/Llama-3.2-90B-Vision | Not documented | — | — | — | |||
| meta-llama/Llama-3.2-11B-Visionmeta-llama/Llama-3.2-11B-Vision | Not documented | — | — | — | |||
| meta-llama/Llama-Guard-3-1Bmeta-llama/Llama-Guard-3-1B | Not documented | — | — | — | |||
| ai21.j2-ultra-v1bedrock/ai21.j2-ultra-v1 | 8.191K | $18.8 | $18.8 | — | |||
| ai21.jamba-1-5-large-v1:0bedrock/ai21.jamba-1-5-large-v1:0 | 256K | $2 | $8 | — | |||
| ai21.jamba-1-5-mini-v1:0bedrock/ai21.jamba-1-5-mini-v1:0 | 256K | $0.2 | $0.4 | — | |||
| microsoft/Mage-VLmicrosoft/Mage-VL | Not documented | — | — | — | |||
| ai21.jamba-instruct-v1:0bedrock/ai21.jamba-instruct-v1:0 | 70K | $0.5 | $0.7 | — | |||
| meta-llama/Llama-Guard-3-1B-INT4meta-llama/Llama-Guard-3-1B-INT4 | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-405B-Instructmeta-llama/Llama-3.1-405B-Instruct | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-405Bmeta-llama/Llama-3.1-405B | Not documented | — | — | — | |||
| google/gemma-4-31B-itdeepinfra/google/gemma-4-31b-it | 262.144K | $0.13 | $0.38 | — | |||
| us.writer.palmyra-x4-v1:0bedrock_converse/us.writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| us.writer.palmyra-x5-v1:0bedrock_converse/us.writer.palmyra-x5-v1:0 | 1M | $0.6 | $6 | — | |||
| databricks-gpt-5-1-codex-maxdatabricks/databricks-gpt-5-1-codex-max | 272K | $1.25 | $10 | — | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 | $5 | — | |||
| microsoft/Mage-Flow-Turbomicrosoft/Mage-Flow-Turbo | Not documented | — | — | — | |||
| meta-llama/Llama-3.1-8B-Instructmeta-llama/Llama-3.1-8B-Instruct | Not documented | — | — | — | |||
| inclusionai/ling-3.0-flashnovita/inclusionai/ling-3.0-flash | 262.144K | $0.06 | $0.18 | — |