No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
No provider description is available for this model yet.
The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| grok-4-fast-non-reasoningazure_ai/grok-4-fast-non-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| grok-4-fast-reasoningazure_ai/grok-4-fast-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| grok-4-1-fast-non-reasoningazure_ai/grok-4-1-fast-non-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| nvidia/MiniMax-M2.7-DFlashnvidia/MiniMax-M2.7-DFlash | Not documented | — | — | — | |||
| grok-code-fast-1azure_ai/grok-code-fast-1 | 131.072K | $0.2 | $1.5 | — | |||
| jais-30b-chatazure_ai/jais-30b-chat | 8.192K | $3200 | $9710 | — | |||
| jamba-instructazure_ai/jamba-instruct | 70K | $0.5 | $0.7 | — | |||
| nvidia/Nemotron-Labs-Audex-30B-A3Bnvidia/Nemotron-Labs-Audex-30B-A3B | Not documented | — | — | — | |||
| kimi-k2.6azure_ai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| ministral-3bazure_ai/ministral-3b | 128K | $0.04 | $0.04 | — | |||
| mistral-largeazure_ai/mistral-large | 32K | $4 | $12 | — | |||
| us.twelvelabs.pegasus-1-2-v1:0bedrock/us.twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| OpenAI: GPT Chat Latestopenai/gpt-chat-latest | 400K | $5 | $30 | — | |||
| mistral-large-latestazure_ai/mistral-large-latest | 128K | $2 | $6 | — | |||
| mistral-large-3azure_ai/mistral-large-3 | 256K | $0.5 | $1.5 | — | |||
| mistral-medium-2505azure_ai/mistral-medium-2505 | 131.072K | $0.4 | $2 | — | |||
| mistral-nemoazure_ai/mistral-nemo | 131.072K | $0.15 | $0.15 | — | |||
| mistral-smallazure_ai/mistral-small | 32K | $1 | $3 | — | |||
| nvidia/Kimi-K2.7-Code-DFlashnvidia/Kimi-K2.7-Code-DFlash | Not documented | — | — | — | |||
| nvidia/Cosmos3-Super-Text2Imagenvidia/Cosmos3-Super-Text2Image | Not documented | — | — | — | |||
| babbage-002text-completion-openai/babbage-002 | 16.384K | $0.4 | $0.4 | — | |||
| */1-month-commitment/cohere.command-light-text-v14bedrock/*/1-month-commitment/cohere.command-light-text-v14 | 4.096K | — | — | — | |||
| */1-month-commitment/cohere.command-text-v14bedrock/*/1-month-commitment/cohere.command-text-v14 | 4.096K | — | — | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1M | $3 | $15 | — | |||
| */6-month-commitment/cohere.command-light-text-v14bedrock/*/6-month-commitment/cohere.command-light-text-v14 | 4.096K | — | — | — | |||
| twelvelabs.pegasus-1-2-v1:0bedrock/twelvelabs.pegasus-1-2-v1:0 | Not documented | — | $7.5 | — | |||
| ap-northeast-1/1-month-commitment/anthropic.claude-instant-v1bedrock/ap-northeast-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| nvidia/Nemotron-Labs-Audex-2Bnvidia/Nemotron-Labs-Audex-2B | Not documented | — | — | — | |||
| ap-northeast-1/1-month-commitment/anthropic.claude-v2:1bedrock/ap-northeast-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| ap-northeast-1/6-month-commitment/anthropic.claude-instant-v1bedrock/ap-northeast-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| global-standard/gpt-4o-miniazure/global-standard/gpt-4o-mini | 128K | $0.15 | $0.6 | — | |||
| ap-northeast-1/6-month-commitment/anthropic.claude-v1bedrock/ap-northeast-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| Qwen: Qwen3 235B A22Bqwen/qwen3-235b-a22b | 131.072K | $0.455 | $1.82 | — | |||
| ap-northeast-1/6-month-commitment/anthropic.claude-v2:1bedrock/ap-northeast-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| ap-northeast-1/anthropic.claude-instant-v1bedrock/ap-northeast-1/anthropic.claude-instant-v1 | 100K | $2.23 | $7.55 | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| ap-northeast-1/anthropic.claude-v2:1bedrock/ap-northeast-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| ap-northeast-1/deepseek.v3.2bedrock/ap-northeast-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| ap-northeast-1/minimax.minimax-m2.1bedrock/ap-northeast-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| ap-northeast-1/minimax.minimax-m2.5bedrock/ap-northeast-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| ap-northeast-1/moonshotai.kimi-k2-thinkingbedrock/ap-northeast-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| OpenAI: GPT-4 Turbo (batch)openai/gpt-4-turbo:batch | 128K | $5 | $15 | — | |||
| MoonshotAI: Kimi K2 0711moonshotai/kimi-k2 | 131.072K | $0.57 | $2.3 | — | |||
| ap-northeast-1/qwen.qwen3-coder-nextbedrock/ap-northeast-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| moonshotai.kimi-k2-thinkingbedrock/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| moonshotai.kimi-k2.5bedrock/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3.03 | — | |||
| ap-south-1/meta.llama3-70b-instruct-v1:0bedrock/ap-south-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $3.18 | $4.2 | — | |||
| nvidia/Nemotron-Cascade-2-30B-A3Bnvidia/Nemotron-Cascade-2-30B-A3B | Not documented | — | — | — |