No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic...
GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| deepseek-ai/DeepSeek-V4-Flash-0731nebius/deepseek-ai/deepseek-v4-flash-0731 | 1.024M | $0.14 | $0.28 | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v2:1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| Z.ai: GLM 4.6Vz-ai/glm-4.6v | 131.072K | $0.3 | $0.9 | — | |||
| Deep Cogito: Cogito v2.1 671Bdeepcogito/cogito-v2.1-671b | 128K | $1.25 | $1.25 | — | |||
| deepseek-ai/DeepSeek-V4-Flashnebius/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| qwen3-max-previewqwen_ai_platform/qwen3-max-preview | 258.048K | — | — | — | |||
| openai/gpt-oss-120b-Ultradeepinfra/openai/gpt-oss-120b-ultra | 131.072K | $0.2 | $0.95 | — | |||
| Qwen/Qwen3.8-Maxdeepinfra/qwen/qwen3.8-max | 256K | $1.65 | $4.951 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| Relace: Relace Searchrelace/relace-search | 256K | $1 | $3 | — | |||
| OpenAI: GPT-5.2 Chatopenai/gpt-5.2-chat | 128K | $1.75 | $14 | — | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 202.752K | $0.4 | $1.75 | — | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 204.8K | $0.3 | $1.2 | — | |||
| databricks-qwen35-122b-a10bdatabricks/databricks-qwen35-122b-a10b | 262.144K | $0.22 | $2.2 | — | |||
| eu-central-1/6-month-commitment/anthropic.claude-v1bedrock/eu-central-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| qwen3-coder-plus-2025-07-22qwen_ai_platform/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| eu-central-1/anthropic.claude-instant-v1bedrock/eu-central-1/anthropic.claude-instant-v1 | 100K | $2.48 | $8.38 | — | |||
| eu-central-1/anthropic.claude-v1bedrock/eu-central-1/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| eu-central-1/anthropic.claude-v2:1bedrock/eu-central-1/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| OpenAI: GPT Audioopenai/gpt-audio | 128K | $2.5 | $10 | — | |||
| MiniMax: MiniMax M2-herminimax/minimax-m2-her | 65.536K | $0.3 | $1.2 | — | |||
| eu-central-1/minimax.minimax-m2.1bedrock/eu-central-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-central-1/minimax.minimax-m2.5bedrock/eu-central-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| eu-west-1/meta.llama3-70b-instruct-v1:0bedrock/eu-west-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.86 | $3.78 | — | |||
| eu-west-1/meta.llama3-8b-instruct-v1:0bedrock/eu-west-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.32 | $0.65 | — | |||
| eu-west-1/minimax.minimax-m2.1bedrock/eu-west-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| eu-west-1/minimax.minimax-m2.5bedrock/eu-west-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| Upstage: Solar Pro 3upstage/solar-pro-3 | 131.072K | $0.15 | $0.6 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| SpaceXAI: Grok 4.5x-ai/grok-4.5 | 500K | $2 | $6 | — | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1M | $0.26 | $1.56 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1M | $1 | $2 | — | |||
| Inception: Mercury 2.5inception/mercury-2.5 | 260K | $0.04 | $0.15 | — | |||
| databricks-grok-4-6databricks/databricks-grok-4-6 | 500K | $2.5 | $7.5 | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| eu-west-1/qwen.qwen3-coder-nextbedrock/eu-west-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| eu-west-2/meta.llama3-70b-instruct-v1:0bedrock/eu-west-2/meta.llama3-70b-instruct-v1:0 | 8.192K | $3.45 | $4.55 | — | |||
| eu-west-2/meta.llama3-8b-instruct-v1:0bedrock/eu-west-2/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.39 | $0.78 | — | |||
| AionLabs: Aion-2.0aion-labs/aion-2.0 | 131.072K | $0.8 | $1.6 | — | |||
| eu-west-2/minimax.minimax-m2.5bedrock/eu-west-2/minimax.minimax-m2.5 | 1M | $0.47 | $1.86 | — | |||
| eu-west-2/qwen.qwen3-coder-nextbedrock/eu-west-2/qwen.qwen3-coder-next | 262.144K | $0.78 | $1.86 | — | |||
| eu-west-3/mistral.mistral-7b-instruct-v0:2bedrock/eu-west-3/mistral.mistral-7b-instruct-v0:2 | 32K | $0.2 | $0.26 | — | |||
| Qwen: Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b | 262.144K | $0.26 | $2.08 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| eu-west-3/mistral.mistral-large-2402-v1:0bedrock/eu-west-3/mistral.mistral-large-2402-v1:0 | 32K | $10.4 | $31.2 | — | |||
| databricks-gpt-5-6-lunadatabricks/databricks-gpt-5-6-luna | 922K | $1 | $6 | — |