No provider description is available for this model yet.

azure_ai/mistral-large-latest 128K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/mistral-large-3 256K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/mistral-medium-2505 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

azure_ai/mistral-nemo 131.072K context $0.15/M input $0.15/M output

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

anthropic/claude-fable-5:batch 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/mistral-small-2503 128K context $0.1/M input $0.3/M output

No provider description is available for this model yet.

anthropic/claude-3-opus-20240229 200K context $15/M input $75/M output

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:batch 524.288K context $1/M input $4.05/M output

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

minimax/minimax-m2.7 204.8K context $0.3/M input $1.2/M output

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2 128K context $0.25/M input $0.75/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5:batch 200K context $0.5/M input $2.5/M output

No provider description is available for this model yet.

bedrock/moonshotai.kimi-k2-thinking 262.144K context $0.73/M input $3.03/M output

No provider description is available for this model yet.

bedrock/moonshotai.kimi-k2.5 262.144K context $0.6/M input $3.03/M output

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...

qwen/qwen3-max 262.144K context $0.78/M input $3.9/M output

No provider description is available for this model yet.

bedrock/ap-south-1/deepseek.v3.2 163.84K context $0.74/M input $2.22/M output

No provider description is available for this model yet.

cloudflare/@cf/openai/gpt-oss-20b 128K context $0.2/M input $0.3/M output

No provider description is available for this model yet.

bedrock/cohere.command-r-v1:0 128K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

cohere_chat/command-a-03-2025 256K context $2.5/M input $10/M output

No provider description is available for this model yet.

cohere_chat/command-r 128K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

cohere_chat/command-r-08-2024 128K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

cohere_chat/command-r-plus 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/us/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

cohere_chat/command-r7b-12-2024 128K context $0.037/M input $0.15/M output

No provider description is available for this model yet.

dashscope/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

dashscope/qwen-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen-flash-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen-plus 129.024K context $0.4/M input $1.2/M output

No provider description is available for this model yet.

dashscope/qwen-plus-2025-01-25 129.024K context $0.4/M input $1.2/M output

No provider description is available for this model yet.

dashscope/qwen-plus-2025-04-28 129.024K context $0.4/M input $1.2/M output

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

z-ai/glm-5.1 200K context $0.966/M input $3.036/M output

No provider description is available for this model yet.

dashscope/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

azure/eu/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output