735 models

No provider description is available for this model yet.

azure/us/gpt-5.5 1.05M context $5.5/M input $33/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

azure/gpt-5.5 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-pro 1M context $1.74/M input $3.48/M output

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol:batch 1.05M context $1/M input $5/M output

No provider description is available for this model yet.

azure_ai/deepseek-v4-flash 1M context $0.19/M input $0.51/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7:batch 1M context $2.5/M input $12.5/M output

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5:batch 1M context $1/M input $5/M output

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

anthropic/claude-fable-5:batch 1M context $5/M input $25/M output

No provider description is available for this model yet.

azure/eu/gpt-5.6-luna 1.05M context $1.1/M input $6.6/M output

No provider description is available for this model yet.

dashscope/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

dashscope/qwen-turbo-2024-11-01 1M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

dashscope/qwen-turbo-2025-04-28 1M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

dashscope/qwen-turbo-latest 1M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

fireworks_ai/deepseek-v4-flash 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

fireworks_ai/deepseek-v4-pro 1.04858M context $1.74/M input $3.48/M output

No provider description is available for this model yet.

fireworks_ai/glm-5p2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openai/ft:gpt-4.1-mini-2025-04-14 1.04758M context $0.8/M input $3.2/M output

No provider description is available for this model yet.

openai/ft:gpt-4.1-nano-2025-04-14 1.04758M context $0.2/M input $0.8/M output

This model always redirects to the latest model in the OpenAI GPT family.

~openai/gpt-latest 1.05M context $2/M input $10/M output

No provider description is available for this model yet.

vertex_ai/gemini-3-pro-preview 1.04858M context $2/M input $12/M output

Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

anthropic/claude-opus-4.7-fast 1M context $30/M input $150/M output