No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen/Qwen3.8-Flashtogether_ai/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| OpenAI: GPT-5.6 Terra (batch)openai/gpt-5.6-terra:batch | 1.05M | $1 | $6 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1M | $1.5 | $7.5 | — | |||
| Anthropic: Claude Sonnet 5 (batch)anthropic/claude-sonnet-5:batch | 1M | $1 | $5 | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1.04758M | $1 | $4 | — | |||
| Qwen: Qwen3.8 2.4T A95Bqwen/qwen3.8-2.4t-a95b | 1M | $2 | $6 | — | |||
| Anthropic: Claude Opus 4.7 (batch)anthropic/claude-opus-4.7:batch | 1M | $2.5 | $12.5 | — | |||
| Qwen: Qwen3.6 Flashqwen/qwen3.6-flash | 1M | $0.188 | $1.125 | — | |||
| us.openai.gpt-6-astrabedrock_converse/us.openai.gpt-6-astra | 1.05M | $11 | $55 | — | |||
| global.openai.gpt-6-astrabedrock_converse/global.openai.gpt-6-astra | 1.05M | $10 | $50 | — | |||
| lyria-3.5gemini/lyria-3.5 | 1.04858M | — | — | — | |||
| databricks-claude-fable-5-1databricks/databricks-claude-fable-5-1 | 1M | $10 | $50 | — | |||
| databricks-gemini-3-8-flashdatabricks/databricks-gemini-3-8-flash | 1.04858M | — | — | — | |||
| databricks-gemini-3-7-flashdatabricks/databricks-gemini-3-7-flash | 1.04858M | — | — | — | |||
| databricks-gemini-3-6-flashdatabricks/databricks-gemini-3-6-flash | 1.04858M | $1.875 | $9.375 | — | |||
| databricks-gemini-3-5-flashdatabricks/databricks-gemini-3-5-flash | 1.04858M | $1.875 | $11.25 | — | |||
| databricks-gemini-3-5-flash-litedatabricks/databricks-gemini-3-5-flash-lite | 1.04858M | $0.375 | $3.125 | — | |||
| databricks-glm-5-3databricks/databricks-glm-5-3 | 1.04858M | $1.4 | $4.4 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| deepseek-ai/DeepSeek-V4-Flashnebius/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731nebius/deepseek-ai/deepseek-v4-flash-0731 | 1.024M | $0.14 | $0.28 | — | |||
| deepseek-ai/DeepSeek-V4-Pronebius/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.75 | $3.5 | — | |||
| MiniMaxAI/MiniMax-M3nebius/minimaxai/minimax-m3 | 1.04858M | $0.3 | $1.2 | — | |||
| moonshotai/Kimi-K3nebius/moonshotai/kimi-k3 | 1.024M | $3 | $15 | — | |||
| nvidia/Nemotron-3-Ultra-550b-a55bnebius/nvidia/nemotron-3-ultra-550b-a55b | 1.04858M | $1 | $3 | — | |||
| nvidia/Nemotron-3_5-Lightningnebius/nvidia/nemotron-3_5-lightning | 1.04858M | $0.06 | $0.24 | — | |||
| zai-org/GLM-5.2nebius/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| zai-org/GLM-5.3-Flashnebius/zai-org/glm-5.3-flash | 1.024M | $0.15 | $0.5 | — | |||
| claude-mythos-5-1anthropic/claude-mythos-5-1 | 1M | $10 | $50 | — | |||
| accounts/fireworks/models/deepseek-v4-flash-vision-expfireworks_ai/accounts/fireworks/models/deepseek-v4-flash-vision-exp | 1.04858M | $0.22 | $0.66 | — | |||
| deepseek-v4-flash-vision-expfireworks_ai/deepseek-v4-flash-vision-exp | 1.04858M | $0.22 | $0.66 | — | |||
| anthropic/claude-fable-5openrouter/anthropic/claude-fable-5 | 1M | $10 | $50 | — | |||
| anthropic/claude-fable-5.1openrouter/anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| anthropic/claude-opus-4.8openrouter/anthropic/claude-opus-4.8 | 1M | $5 | $25 | — | |||
| anthropic/claude-sonnet-5openrouter/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| google/gemini-2.5-flash-liteopenrouter/google/gemini-2.5-flash-lite | 1.04858M | $0.1 | $0.4 | — | |||
| google/gemini-3.5-flashopenrouter/google/gemini-3.5-flash | 1.04858M | $1.5 | $9 | — | |||
| google/gemini-3.5-flash-liteopenrouter/google/gemini-3.5-flash-lite | 1.04858M | $0.3 | $2.5 | — | |||
| google/gemini-3.6-flashopenrouter/google/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | — | |||
| google/gemini-3.7-flashopenrouter/google/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| google/gemini-3.8-flashopenrouter/google/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| openai/gpt-5.4openrouter/openai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| openai/gpt-5.5openrouter/openai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| x-ai/grok-4.20openrouter/x-ai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| x-ai/grok-4.20-multi-agentopenrouter/x-ai/grok-4.20-multi-agent | 1M | $1.25 | $2.5 | — | |||
| x-ai/grok-4.3openrouter/x-ai/grok-4.3 | 1M | $1.25 | $2.5 | — | |||
| zai-org/GLM-5.3baseten/zai-org/glm-5.3 | 1.04858M | $1.4 | $4.4 | — |