735 models

No provider description is available for this model yet.

together_ai/qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

openai/gpt-5.6-terra:batch 1.05M context $1/M input $6/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3:free 1.04858M context Free input Free output

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

anthropic/claude-sonnet-4.5:batch 1M context $1.5/M input $7.5/M output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash:batch 1.04858M context $0.15/M input $1.25/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview-05-06 1.04858M context $1.25/M input $10/M output

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

openai/gpt-4.1:batch 1.04758M context $1/M input $4/M output

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

qwen/qwen3.8-2.4t-a95b 1M context $2/M input $6/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7:batch 1M context $2.5/M input $12.5/M output

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

qwen/qwen3.6-flash 1M context $0.188/M input $1.125/M output

No provider description is available for this model yet.

gemini/lyria-3.5 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

databricks/databricks-glm-5-3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

databricks/databricks-inkling 1M context $1/M input $4.05/M output

No provider description is available for this model yet.

nebius/deepseek-ai/deepseek-v4-pro 1.04858M context $1.75/M input $3.5/M output

No provider description is available for this model yet.

nebius/minimaxai/minimax-m3 1.04858M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k3 1.024M context $3/M input $15/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

nebius/zai-org/glm-5.3-flash 1.024M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

anthropic/claude-mythos-5-1 1M context $10/M input $50/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.5-flash 1.04858M context $1.5/M input $9/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.6-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

openrouter/google/gemini-3.8-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.20 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

openrouter/x-ai/grok-4.3 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

baseten/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output