No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| mistral-large-latestazure_ai/mistral-large-latest | 128K | $2 | $6 | — | |||
| mistral-large-3azure_ai/mistral-large-3 | 256K | $0.5 | $1.5 | — | |||
| mistral-medium-2505azure_ai/mistral-medium-2505 | 131.072K | $0.4 | $2 | — | |||
| mistral-nemoazure_ai/mistral-nemo | 131.072K | $0.15 | $0.15 | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| mistral-small-2503azure_ai/mistral-small-2503 | 128K | $0.1 | $0.3 | — | |||
| claude-3-opus-20240229anthropic/claude-3-opus-20240229 | 200K | $15 | $75 | — | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 524.288K | $1 | $4.05 | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| ap-northeast-1/deepseek.v3.2bedrock/ap-northeast-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| ap-northeast-1/minimax.minimax-m2.1bedrock/ap-northeast-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Inception: Mercury 2inception/mercury-2 | 128K | $0.25 | $0.75 | — | |||
| ap-northeast-1/minimax.minimax-m2.5bedrock/ap-northeast-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| ap-northeast-1/moonshotai.kimi-k2-thinkingbedrock/ap-northeast-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 200K | $0.5 | $2.5 | — | |||
| ap-northeast-1/moonshotai.kimi-k2.5bedrock/ap-northeast-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| ap-northeast-1/qwen.qwen3-coder-nextbedrock/ap-northeast-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| moonshotai.kimi-k2-thinkingbedrock/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| moonshotai.kimi-k2.5bedrock/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3.03 | — | |||
| Qwen: Qwen3 Maxqwen/qwen3-max | 262.144K | $0.78 | $3.9 | — | |||
| ap-south-1/deepseek.v3.2bedrock/ap-south-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| ap-south-1/minimax.minimax-m2.1bedrock/ap-south-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| @cf/aisingapore/gemma-sea-lion-v4-27b-itcloudflare/@cf/aisingapore/gemma-sea-lion-v4-27b-it | 128K | $0.351 | $0.555 | — | |||
| claude-4-sonnet-20250514anthropic/claude-4-sonnet-20250514 | 1M | $3 | $15 | — | |||
| claude-sonnet-4-5-20250929-v1:0bedrock/claude-sonnet-4-5-20250929-v1:0 | 200K | $3 | $15 | — | |||
| claude-opus-4-6-20260205anthropic/claude-opus-4-6-20260205 | 1M | $5 | $25 | — | |||
| claude-opus-4-7-20260416anthropic/claude-opus-4-7-20260416 | 1M | $5 | $25 | — | |||
| @cf/google/gemma-4-26b-a4b-itcloudflare/@cf/google/gemma-4-26b-a4b-it | 256K | $0.1 | $0.3 | — | |||
| @cf/mistralai/mistral-small-3.1-24b-instructcloudflare/@cf/mistralai/mistral-small-3.1-24b-instruct | 128K | $0.351 | $0.555 | — | |||
| @cf/meta/llama-3.2-11b-vision-instructcloudflare/@cf/meta/llama-3.2-11b-vision-instruct | 128K | $0.048 | $0.676 | — | |||
| @cf/openai/gpt-oss-20bcloudflare/@cf/openai/gpt-oss-20b | 128K | $0.2 | $0.3 | — | |||
| @cf/meta/llama-4-scout-17b-16e-instructcloudflare/@cf/meta/llama-4-scout-17b-16e-instruct | 131K | $0.27 | $0.85 | — | |||
| cohere.command-r-plus-v1:0bedrock/cohere.command-r-plus-v1:0 | 128K | $3 | $15 | — | |||
| cohere.command-r-v1:0bedrock/cohere.command-r-v1:0 | 128K | $0.5 | $1.5 | — | |||
| command-a-03-2025cohere_chat/command-a-03-2025 | 256K | $2.5 | $10 | — | |||
| command-rcohere_chat/command-r | 128K | $0.15 | $0.6 | — | |||
| command-r-08-2024cohere_chat/command-r-08-2024 | 128K | $0.15 | $0.6 | — | |||
| command-r-pluscohere_chat/command-r-plus | 128K | $2.5 | $10 | — | |||
| us/gpt-4o-mini-2024-07-18azure/us/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| command-r7b-12-2024cohere_chat/command-r7b-12-2024 | 128K | $0.037 | $0.15 | — | |||
| qwen-coderdashscope/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashdashscope/qwen-flash | 997.952K | — | — | — | |||
| qwen-flash-2025-07-28dashscope/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plusdashscope/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25dashscope/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28dashscope/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| Z.ai: GLM 5.1z-ai/glm-5.1 | 200K | $0.966 | $3.036 | — | |||
| qwen-plus-2025-07-28dashscope/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11dashscope/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| eu/gpt-5-mini-2025-08-07azure/eu/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — |