1,424 models

No provider description is available for this model yet.

azure_ai/kimi-k2.5 262.144K context $0.6/M input $3/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7:batch 1M context $2.5/M input $12.5/M output

No provider description is available for this model yet.

azure/gpt-4o 128K context $2.5/M input $10/M output

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

qwen/qwen3.8-27b 262.144K context $0.214/M input $2.55/M output

No provider description is available for this model yet.

azure/gpt-4.5-preview 128K context $75/M input $150/M output

No provider description is available for this model yet.

xai/grok-build-latest 500K context $2/M input $6/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1:batch 200K context $7.5/M input $37.5/M output

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol:batch 1.05M context $1/M input $5/M output

No provider description is available for this model yet.

azure/gpt-4.1-nano-2025-04-14 1.04758M context $0.1/M input $0.4/M output

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini:batch 400K context $0.125/M input $1/M output

No provider description is available for this model yet.

azure_ai/phi-4-multimodal-instruct 131.072K context $0.08/M input $0.32/M output

No provider description is available for this model yet.

azure/gpt-4.1-nano 1.04758M context $0.1/M input $0.4/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

azure_ai/phi-3.5-vision-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure/gpt-4.1-mini-2025-04-14 1.04758M context $0.4/M input $1.6/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano:batch 1.04758M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

azure/gpt-4.1-mini 1.04758M context $0.4/M input $1.6/M output

No provider description is available for this model yet.

azure/gpt-4.1-2025-04-14 1.04758M context $2/M input $8/M output

No provider description is available for this model yet.

azure_ai/gpt-5.5-2026-04-23 1.05M context $5/M input $30/M output

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

anthropic/claude-opus-4.6:batch 1M context $2.5/M input $12.5/M output

No provider description is available for this model yet.

azure/us/o4-mini-2025-04-16 200K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/us/o3-2025-04-16 200K context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/gpt-4.1 1.04758M context $2/M input $8/M output

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

anthropic/claude-fable-5.1 1M context $10/M input $50/M output

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3:batch 200K context $1/M input $4/M output

No provider description is available for this model yet.

azure/gpt-4-turbo-vision-preview 128K context $10/M input $30/M output

No provider description is available for this model yet.

azure/us/o1-2024-12-17 200K context $16.5/M input $66/M output

No provider description is available for this model yet.

azure/us/gpt-5.1-chat 128K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5.1 272K context $1.38/M input $11/M output

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

google/gemini-3.1-pro-preview:batch 1.04858M context $1/M input $6/M output

No provider description is available for this model yet.

azure/us/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

No provider description is available for this model yet.

azure/gpt-4-turbo-2024-04-09 128K context $10/M input $30/M output