No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
No provider description is available for this model yet.
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
No provider description is available for this model yet.
This model always redirects to the latest model in the Claude Haiku family.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
No provider description is available for this model yet.
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
No provider description is available for this model yet.
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
This model always redirects to the latest Grok model from xAI.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
No provider description is available for this model yet.
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| nvidia/Ising-Calibration-1-35B-A3Bnvidia/Ising-Calibration-1-35B-A3B | Not documented | — | — | — | |||
| OpenAI: o1-pro (batch)openai/o1-pro:batch | 200K | $75 | $300 | — | |||
| gemini-robotics-er-1.6-previewgemini/gemini-robotics-er-1.6-preview | 131.072K | $1 | $5 | — | |||
| Qwen: Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b | 262.144K | $0.26 | $2.08 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 | $0.75 | — | |||
| gemini-robotics-er-2-previewgemini/gemini-robotics-er-2-preview | 131.072K | $2 | $10 | — | |||
| Anthropic: Claude Haiku Latest~anthropic/claude-haiku-latest | 200K | $1 | $5 | — | |||
| us-gov-east-1/amazon.nova-pro-v1:0bedrock/us-gov-east-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1 | $5 | — | |||
| Z.ai: GLM 5.3 Flashz-ai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| openai/gpt-4o-minigmi/openai/gpt-4o-mini | 131.072K | $0.15 | $0.6 | — | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1.04758M | $1 | $4 | — | |||
| anthropic.claude-haiku-4-5@20251001bedrock_converse/anthropic.claude-haiku-4-5@20251001 | 200K | $1 | $5 | — | |||
| inclusionAI: Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl | 131.072K | $0.06 | $0.18 | — | |||
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 256K | Free | Free | — | |||
| anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/anthropic.claude-3-5-sonnet-20240620-v1:0 | 1M | $3 | $15 | — | |||
| Thinking Machines: Inkling Small (batch)thinkingmachines/inkling-small:batch | 524.288K | $0.5 | $1.2 | — | |||
| Qwen: Qwen3.5-27Bqwen/qwen3.5-27b | 262.144K | $0.195 | $1.56 | — | |||
| xAI: Grok Latest~x-ai/grok-latest | 500K | $2 | $6 | — | |||
| Google: Gemini 3.7 Flash (batch)google/gemini-3.7-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| anthropic.claude-3-7-sonnet-20250219-v1:0bedrock_converse/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3 | $15 | — | |||
| Meta: Llama Guard 4 12Bmeta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| amazon.nova-pro-v1:0bedrock_converse/amazon.nova-pro-v1:0 | 300K | $0.8 | $3.2 | — | |||
| us.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/us.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| anthropic.claude-3-sonnet-20240229-v1:0bedrock/anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1M | $1.5 | $7.5 | — | |||
| anthropic.claude-opus-4-1-20250805-v1:0bedrock_converse/anthropic.claude-opus-4-1-20250805-v1:0 | 200K | $15 | $75 | — | |||
| us.amazon.nova-2-lite-v1:0bedrock_converse/us.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| anthropic.claude-opus-4-5-20251101-v1:0bedrock_converse/anthropic.claude-opus-4-5-20251101-v1:0 | 200K | $5 | $25 | — | |||
| Meta: Llama 3.2 11B Vision Instructmeta-llama/llama-3.2-11b-vision-instruct | 131.072K | $0.345 | $0.345 | — | |||
| anthropic.claude-opus-4-6-v1bedrock_converse/anthropic.claude-opus-4-6-v1 | 1M | $5 | $25 | — | |||
| global.anthropic.claude-opus-4-6-v1bedrock_converse/global.anthropic.claude-opus-4-6-v1 | 1M | $5 | $25 | — | |||
| us.anthropic.claude-opus-4-6-v1bedrock_converse/us.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| eu.anthropic.claude-opus-4-6-v1bedrock_converse/eu.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| OpenAI: GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch | 1.05M | $15 | $90 | — | |||
| google/gemini-3-pro-previewgmi/google/gemini-3-pro-preview | 1.04858M | $2 | $12 | — | |||
| google/gemini-3-flash-previewgmi/google/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | — | |||
| au.anthropic.claude-opus-4-6-v1bedrock_converse/au.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| eu.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/eu.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| anthropic.claude-opus-4-7bedrock_converse/anthropic.claude-opus-4-7 | 1M | $5 | $25 | — | |||
| anthropic.claude-mythos-previewbedrock/anthropic.claude-mythos-preview | 1M | — | — | — | |||
| us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch | 1.05M | $0.1 | $0.6 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 400K | $0.625 | $5 | — | |||
| chatgpt-4o-latestopenai/chatgpt-4o-latest | 128K | $5 | $15 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 200K | $0.5 | $2.5 | — |