No provider description is available for this model yet.

azure_ai/claude-opus-4-6 1M context $5/M input $25/M output

No provider description is available for this model yet.

deepinfra/openai/gpt-oss-120b-ultra 131.072K context $0.2/M input $0.95/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-7 1M context $5/M input $25/M output

No provider description is available for this model yet.

openrouter/openai/gpt-6-astra-pro 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

azure_ai/claude-fable-5 1M context $10/M input $50/M output

No provider description is available for this model yet.

azure_ai/claude-opus-5 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-8 1M context $5/M input $25/M output

Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...

inflection/inflection-3-productivity 8K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure_ai/claude-opus-4-1 200K context $15/M input $75/M output

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

sakana/namazu 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/snowflake-llama-3.1-405b 8K context Input not listed Output not listed

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b:batch 1.01M context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-5 1M context $2/M input $10/M output

No provider description is available for this model yet.

azure_ai/claude-sonnet-4-6 1M context $3/M input $15/M output

No provider description is available for this model yet.

azure/computer-use-preview 8.192K context $3/M input $12/M output

No provider description is available for this model yet.

azure/container Not documented context Input not listed Output not listed

No provider description is available for this model yet.

vertex_ai/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

azure_ai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

openai/gpt-3.5-turbo:batch 16.385K context $0.25/M input $0.75/M output

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

moonshotai/kimi-k2.7-code:batch 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

gemini/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

x-ai/grok-4.3:batch 1M context $1/M input $2/M output

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

z-ai/glm-4.5v 65.536K context $0.6/M input $1.8/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

xai/grok-build-latest 500K context $2/M input $6/M output

No provider description is available for this model yet.

scaleway/glm-5.2 256K context $1.8/M input $5.5/M output

No provider description is available for this model yet.

gradient_ai/llama3.3-70b-instruct 128K context $0.65/M input $0.65/M output