No provider description is available for this model yet.

azure/o4-mini 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/o4-mini-2025-04-16 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-2025-04-14 1.04758M context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-mini-2025-04-14 1.04758M context $0.44/M input $1.76/M output

No provider description is available for this model yet.

azure/us/gpt-4.1-nano-2025-04-14 1.04758M context $0.11/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-08-06 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-4o-2024-11-20 128K context $2.75/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-4o-mini-2024-07-18 128K context $0.165/M input $0.66/M output

No provider description is available for this model yet.

azure/us/gpt-5-2025-08-07 272K context $1.375/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5-mini-2025-08-07 272K context $0.275/M input $2.2/M output

No provider description is available for this model yet.

upstage/Llama-2-70b-instruct Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure/us/gpt-5-nano-2025-08-07 272K context $0.055/M input $0.44/M output

No provider description is available for this model yet.

azure/us/gpt-5.1 272K context $1.38/M input $11/M output

No provider description is available for this model yet.

azure/us/gpt-5.1-chat 128K context $1.38/M input $11/M output

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

openai/gpt-oss-20b:batch 131.072K context $0.05/M input $0.2/M output

No provider description is available for this model yet.

azure/us/o1-2024-12-17 200K context $16.5/M input $66/M output

No provider description is available for this model yet.

azure/us/o1-mini-2024-09-12 128K context $1.21/M input $4.84/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1 200K context $15/M input $75/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:free 1.04858M context Free input Free output

No provider description is available for this model yet.

upstage/llama-65b-instruct Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure/us/o1-preview-2024-09-12 128K context $16.5/M input $66/M output

No provider description is available for this model yet.

cerebras/qwen-3-32b 128K context $0.4/M input $0.8/M output

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small:batch 524.288K context $0.5/M input $1.2/M output

No provider description is available for this model yet.

azure/us/o3-2025-04-16 200K context $2.2/M input $8.8/M output

No provider description is available for this model yet.

azure/us/o3-mini-2025-01-31 200K context $1.21/M input $4.84/M output

No provider description is available for this model yet.

azure/us/o4-mini-2025-04-16 200K context $1.21/M input $4.84/M output

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

openai/gpt-4.1:batch 1.04758M context $1/M input $4/M output

No provider description is available for this model yet.

github_copilot/gpt-5 128K context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/llama-3.3-70b-instruct 128K context $0.71/M input $0.71/M output

No provider description is available for this model yet.

azure_ai/meta-llama-3-70b-instruct 8.192K context $1.1/M input $0.37/M output

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano:batch 1.04758M context $0.05/M input $0.2/M output

No provider description is available for this model yet.

azure_ai/phi-3-medium-4k-instruct 4.096K context $0.17/M input $0.68/M output

No provider description is available for this model yet.

upstage/llama-30b-instruct Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/phi-3-mini-128k-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-3-mini-4k-instruct 4.096K context $0.13/M input $0.52/M output