No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| @cf/google/gemma-4-26b-a4b-itcloudflare/@cf/google/gemma-4-26b-a4b-it | 256K | $0.1 | $0.3 | — | |||
| command-a-03-2025cohere_chat/command-a-03-2025 | 256K | $2.5 | $10 | — | |||
| qwen-coderdashscope/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashdashscope/qwen-flash | 997.952K | — | — | — | |||
| qwen-flash-2025-07-28dashscope/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-07-28dashscope/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| OpenAI: o3 (batch)openai/o3:batch | 200K | $1 | $4 | — | |||
| qwen-plus-latestdashscope/qwen-plus-latest | 997.952K | — | — | — | |||
| google/gemini-2.5-prodeepinfra/google/gemini-2.5-pro | 1M | $1.25 | $10 | — | |||
| qwen-turbo-2024-11-01dashscope/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28dashscope/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-latestdashscope/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| claude-opus-4-5azure_ai/claude-opus-4-5 | 200K | $5 | $25 | — | |||
| qwen3-coder-flash-2025-07-28dashscope/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen3-coder-plusdashscope/qwen3-coder-plus | 997.952K | — | — | — | |||
| qwen3-coder-plus-2025-07-22dashscope/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewdashscope/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-maxdashscope/qwen3-max | 258.048K | — | — | — | |||
| qwen3-max-2026-01-23dashscope/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructdashscope/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-next-80b-a3b-thinkingdashscope/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-plusdashscope/qwen3-vl-plus | 260.096K | — | — | — | |||
| qwen3.5-plusdashscope/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxdashscope/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusdashscope/qwen3.7-plus | 991.808K | — | — | — | |||
| databricks-claude-3-7-sonnetdatabricks/databricks-claude-3-7-sonnet | 200K | $3 | $15 | — | |||
| databricks-claude-haiku-4-5databricks/databricks-claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| databricks-claude-opus-4databricks/databricks-claude-opus-4 | 200K | $15 | $75 | — | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 262K | Free | Free | — | |||
| databricks-claude-opus-4-5databricks/databricks-claude-opus-4-5 | 200K | $5 | $25 | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| databricks-claude-sonnet-4-1databricks/databricks-claude-sonnet-4-1 | 200K | $3 | $15 | — | |||
| databricks-claude-sonnet-4-5databricks/databricks-claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| databricks-gemini-2-5-flashdatabricks/databricks-gemini-2-5-flash | 1.04858M | $0.3 | $2.5 | — | |||
| databricks-gemini-2-5-prodatabricks/databricks-gemini-2-5-pro | 1.04858M | $1.25 | $10 | — | |||
| databricks-gpt-5databricks/databricks-gpt-5 | 272K | $1.25 | $10 | — | |||
| databricks-gpt-5-1databricks/databricks-gpt-5-1 | 272K | $1.25 | $10 | — | |||
| databricks-gpt-5-minidatabricks/databricks-gpt-5-mini | 272K | $0.25 | $2 | — | |||
| databricks-gpt-5-nanodatabricks/databricks-gpt-5-nano | 272K | $0.05 | $0.4 | — | |||
| databricks-meta-llama-3-1-8b-instructdatabricks/databricks-meta-llama-3-1-8b-instruct | 200K | $0.15 | $0.45 | — | |||
| Qwen/Qwen3-235B-A22B-Instruct-2507deepinfra/qwen/qwen3-235b-a22b-instruct-2507 | 262.144K | $0.09 | $0.6 | — | |||
| Qwen/Qwen3-235B-A22B-Thinking-2507deepinfra/qwen/qwen3-235b-a22b-thinking-2507 | 262.144K | $0.3 | $2.9 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instructdeepinfra/qwen/qwen3-coder-480b-a35b-instruct | 262.144K | $0.4 | $1.6 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-Turbodeepinfra/qwen/qwen3-coder-480b-a35b-instruct-turbo | 262.144K | $0.29 | $1.2 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Instructdeepinfra/qwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.14 | $1.4 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Thinkingdeepinfra/qwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.14 | $1.4 | — | |||
| us/o1-2024-12-17azure/us/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| anthropic/claude-4-opusdeepinfra/anthropic/claude-4-opus | 200K | $16.5 | $82.5 | — | |||
| anthropic/claude-4-sonnetdeepinfra/anthropic/claude-4-sonnet | 200K | $3.3 | $16.5 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1M | $1 | $2 | — |