3,449 models

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6:batch 1M context $1.5/M input $7.5/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3-0324 128K context $1.14/M input $4.56/M output

No provider description is available for this model yet.

azure_ai/deepseek-v3.1 131.072K context $1.23/M input $4.94/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-flash 1M context $0.19/M input $0.51/M output

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

z-ai/glm-4.5-air 131.072K context $0.13/M input $0.85/M output

No provider description is available for this model yet.

azure_ai/global/grok-3 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/global/grok-3-mini 131.072K context $0.25/M input $1.27/M output

No provider description is available for this model yet.

azure_ai/grok-3 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/grok-3-mini 131.072K context $0.25/M input $1.27/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro:batch 1.04858M context $0.625/M input $5/M output

No provider description is available for this model yet.

azure_ai/grok-4 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/grok-4-fast-non-reasoning 131.072K context $0.2/M input $0.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-fast-reasoning 131.072K context $0.2/M input $0.5/M output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5 1M context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/grok-4-1-fast-reasoning 131.072K context $0.2/M input $0.5/M output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:batch 262.144K context $0.39/M input $0.97/M output

No provider description is available for this model yet.

azure_ai/jais-30b-chat 8.192K context $3200/M input $9710/M output

No provider description is available for this model yet.

azure_ai/jamba-instruct 70K context $0.5/M input $0.7/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.5 262.144K context $0.6/M input $3/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/ministral-3b 128K context $0.04/M input $0.04/M output

No provider description is available for this model yet.

azure_ai/mistral-large 32K context $4/M input $12/M output

No provider description is available for this model yet.

azure_ai/mistral-large-2407 128K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/mistral-large-latest 128K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/mistral-large-3 256K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/mistral-medium-2505 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

azure_ai/mistral-nemo 131.072K context $0.15/M input $0.15/M output

No provider description is available for this model yet.

azure_ai/mistral-small 32K context $1/M input $3/M output

No provider description is available for this model yet.

azure_ai/mistral-small-2503 128K context $0.1/M input $0.3/M output

No provider description is available for this model yet.

text-completion-openai/babbage-002 16.384K context $0.4/M input $0.4/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1:batch 200K context $7.5/M input $37.5/M output

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash:batch 1.04858M context $0.375/M input $1.875/M output

No provider description is available for this model yet.

snowflake/mistral-large 32K context Input not listed Output not listed