No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
No provider description is available for this model yet.
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
No provider description is available for this model yet.
This model always redirects to the latest model in the Claude Fable family.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
No provider description is available for this model yet.
No provider description is available for this model yet.
Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.
No provider description is available for this model yet.
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| glm-5p2-fastfireworks_ai/glm-5p2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| meta-llama/Llama-2-13b-chat-hfmeta-llama/Llama-2-13b-chat-hf | Not documented | — | — | — | |||
| meta-llama/Llama-2-7b-chat-hfmeta-llama/Llama-2-7b-chat-hf | Not documented | — | — | — | |||
| us.amazon.nova-2-lite-v1:0bedrock_converse/us.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| us.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/us.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| amazon.nova-micro-v1:0bedrock_converse/amazon.nova-micro-v1:0 | 128K | $0.035 | $0.14 | — | |||
| deepseek-v4-flash-0731fireworks_ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 131.072K | $0.2 | $1 | — | |||
| amazon.nova-pro-v1:0bedrock_converse/amazon.nova-pro-v1:0 | 300K | $0.8 | $3.2 | — | |||
| OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batch | 400K | $0.125 | $1 | — | |||
| Anthropic: Claude Opus 4.7 (Fast)anthropic/claude-opus-4.7-fast | 1M | $30 | $150 | — | |||
| FW-Inklingazure_ai/fw-inkling | 1.04858M | $1 | $4.05 | — | |||
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| twelvelabs.pegasus-1-2-v1:0bedrock/twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| Perceptron: Perceptron Mk1perceptron/perceptron-mk1 | 32.768K | $0.15 | $1.5 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| google/gemini-3.1-flash-image-previewopenrouter/google/gemini-3.1-flash-image-preview | 65.536K | $0.5 | $3 | — | |||
| Qwen: Qwen3 Coder 480B A35Bqwen/qwen3-coder | 262.144K | $0.3 | $1 | — | |||
| eu.twelvelabs.pegasus-1-2-v1:0bedrock/eu.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| meta-llama/Llama-2-13b-hfmeta-llama/Llama-2-13b-hf | Not documented | — | — | — | |||
| anthropic/claude-fable-5.1openrouter/anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| deepseek-ai/ESFT-gate-math-litedeepseek-ai/ESFT-gate-math-lite | Not documented | — | — | — | |||
| us-gov.nvidia.nemotron-nano-12b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 200K | $15 | $75 | — | |||
| us-gov.nvidia.nemotron-nano-3-30bbedrock_converse/us-gov.nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| meta-llama/Llama-2-7b-hfmeta-llama/Llama-2-7b-hf | Not documented | — | — | — | |||
| amazon.titan-text-lite-v1bedrock/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| amazon.titan-text-premier-v1:0bedrock/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| OpenAI: GPT-3.5 Turbo (batch)openai/gpt-3.5-turbo:batch | 16.385K | $0.25 | $0.75 | — | |||
| deepseek-ai/DeepSeek-V3.2gmi/deepseek-ai/deepseek-v3.2 | 163.84K | $0.28 | $0.4 | — | |||
| Anthropic: Claude Fable Latest~anthropic/claude-fable-latest | 1M | $10 | $50 | — | |||
| anthropic.claude-3-5-haiku-20241022-v1:0bedrock/anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| us-gov-east-1/amazon.nova-pro-v1:0bedrock/us-gov-east-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 | $0.75 | — | |||
| meta-llama/Llama-2-70b-chatmeta-llama/Llama-2-70b-chat | Not documented | — | — | — | |||
| anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1 | $5 | — | |||
| OpenAI: o1-pro (batch)openai/o1-pro:batch | 200K | $75 | $300 | — | |||
| meta-llama/Llama-2-13b-chatmeta-llama/Llama-2-13b-chat | Not documented | — | — | — | |||
| meta-llama/Llama-2-7b-chatmeta-llama/Llama-2-7b-chat | Not documented | — | — | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| FW-GLM-5.2azure_ai/fw-glm-5.2 | 1.04858M | $1.54 | $4.84 | — | |||
| anthropic.claude-haiku-4-5@20251001bedrock_converse/anthropic.claude-haiku-4-5@20251001 | 200K | $1 | $5 | — | |||
| OpenRouter: Fusionopenrouter/fusion | 1M | — | — | — | |||
| anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/anthropic.claude-3-5-sonnet-20240620-v1:0 | 1M | $3 | $15 | — | |||
| openai/gpt-5.4-proopenrouter/openai/gpt-5.4-pro | 1.05M | $30 | $180 | — | |||
| meta-llama/Llama-2-70bmeta-llama/Llama-2-70b | Not documented | — | — | — | |||
| Tencent: Hy-MT2-7Btencent/hy-mt2-7b | 8.192K | $0.074 | $0.295 | — | |||
| us-gov-east-1/amazon.titan-text-lite-v1bedrock/us-gov-east-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 400K | $0.025 | $0.2 | — | |||
| FW-GLM-5.1azure_ai/fw-glm-5.1 | 202.8K | $1.54 | $4.84 | — |