No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| eu.anthropic.claude-opus-4-8bedrock_converse/eu.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| anthropic.claude-opus-4-7bedrock_converse/anthropic.claude-opus-4-7 | 1M | $5 | $25 | — | |||
| zai-org/GLM-4.6baseten/zai-org/glm-4.6 | Not documented | $0.6 | $2.2 | — | |||
| deepseek-ai/DeepSeek-V3.2friendliai/deepseek-ai/deepseek-v3.2 | 163.84K | $0.5 | $1.5 | — | |||
| Baidu: ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b | 123K | $0.42 | $1.25 | — | |||
| google/gemma-4-E2Bgoogle/gemma-4-E2B | Not documented | — | — | — | |||
| zai-org/GLM-4.7baseten/zai-org/glm-4.7 | Not documented | $0.6 | $2.2 | — | |||
| zai-org/GLM-5baseten/zai-org/glm-5 | Not documented | $0.95 | $3.15 | — | |||
| google/gemma-4-E4Bgoogle/gemma-4-E4B | Not documented | — | — | — | |||
| google/gemma-4-26B-A4Bgoogle/gemma-4-26B-A4B | Not documented | — | — | — | |||
| LGAI-EXAONE/K-EXAONE-2.0-750B-A37Bfriendliai/lgai-exaone/k-exaone-2.0-750b-a37b | 262.144K | $0.6 | $2.4 | — | |||
| nvidia/Nemotron-120B-A12Bbaseten/nvidia/nemotron-120b-a12b | Not documented | $0.3 | $0.75 | — | |||
| jp.anthropic.claude-opus-4-8bedrock_converse/jp.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| MiniMaxAI/MiniMax-M2.5baseten/minimaxai/minimax-m2.5 | Not documented | $0.3 | $1.2 | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507 | 262.144K | $0.087 | $0.35 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| anthropic.claude-sonnet-5bedrock_converse/anthropic.claude-sonnet-5 | 1M | $2 | $10 | — | |||
| Qwen: Qwen3.6 27Bqwen/qwen3.6-27b | 262.144K | $0.3 | $2 | — | |||
| Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct | 32.768K | $0.36 | $0.4 | — | |||
| google/gemma-4-31Bgoogle/gemma-4-31B | Not documented | — | — | — | |||
| google/gemma-4-31B-itgoogle/gemma-4-31B-it | Not documented | — | — | — | |||
| MiniMaxAI/MiniMax-M2.1gmi/minimaxai/minimax-m2.1 | 196.608K | $0.3 | $1.2 | — | |||
| eu.anthropic.claude-sonnet-5bedrock_converse/eu.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| au.anthropic.claude-sonnet-5bedrock_converse/au.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| jp.anthropic.claude-sonnet-5bedrock_converse/jp.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| moonshotai/Kimi-K2-Thinkinggmi/moonshotai/kimi-k2-thinking | 262.144K | $0.8 | $1.2 | — | |||
| global.anthropic.claude-sonnet-4-6bedrock_converse/global.anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| zai-org/GLM-5.2friendliai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| us.anthropic.claude-sonnet-4-6bedrock_converse/us.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| au.anthropic.claude-opus-4-6-v1bedrock_converse/au.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| google/gemma-4-31B-itfriendliai/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| OpenAI: GPT-4o Search Previewopenai/gpt-4o-search-preview | 128K | $2.5 | $10 | — | |||
| Qwen: Qwen3 Coder 30B A3B Instructqwen/qwen3-coder-30b-a3b-instruct | 262.144K | $0.07 | $0.28 | — | |||
| eu.anthropic.claude-sonnet-4-6bedrock_converse/eu.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| au.anthropic.claude-sonnet-4-6bedrock_converse/au.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| qwen-3.8-27bcerebras/qwen-3.8-27b | 65.536K | $0.99 | $1.49 | — | |||
| google/gemini-3-flash-previewgmi/google/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | — | |||
| OpenAI: GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch | 1.05M | $15 | $90 | — | |||
| computer-use-previewopenai/computer-use-preview | 8.192K | $3 | $12 | — | |||
| Qwen: Qwen3 Coder Plusqwen/qwen3-coder-plus | 1M | $0.65 | $3.25 | — | |||
| anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-v1bedrock/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| anthropic.claude-v2:1bedrock/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| HuggingFaceH4/zephyr-7b-betaanyscale/huggingfaceh4/zephyr-7b-beta | 16.384K | $0.15 | $0.15 | — | |||
| deepseek/deepseek-v4.1-flashopenrouter/deepseek/deepseek-v4.1-flash | 1.04858M | $0.15 | $0.6 | — | |||
| openai/gpt-5.6-sol-proopenrouter/openai/gpt-5.6-sol-pro | 1.05M | $2 | $10 | — | |||
| google/gemini-3-pro-previewgmi/google/gemini-3-pro-preview | 1.04858M | $2 | $12 | — |