3,646 models

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7 1M context $5/M input $25/M output

No provider description is available for this model yet.

xai/grok-4.20-multi-agent-0309 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

anthropic/claude-mythos-preview 1M context $10/M input $50/M output

No provider description is available for this model yet.

mistral/labs-leanstral-1-5 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

meta-llama/Llama-Guard-4-12B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

meta-llama/Llama-4-Scout-17B-16E Not documented context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-41-copilot Not documented context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o 64K context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/us-east-2/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-05-13 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-08-06 64K context Input not listed Output not listed

This model always redirects to the latest model in the GPT Sol family.

~openai/gpt-sol-latest 1.05M context $2/M input $10/M output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

This model always redirects to the latest model in the GPT Astra family.

~openai/gpt-astra-latest 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

snowflake/llama3.2-3b 128K context Input not listed Output not listed

This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.

openai/gpt-3.5-turbo-instruct 4.095K context $1.5/M input $2/M output

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1:batch 200K context $7.5/M input $30/M output

No provider description is available for this model yet.

snowflake/llama3.3-70b 128K context $0.72/M input $0.72/M output

No provider description is available for this model yet.

snowflake/mistral-7b 32K context Input not listed Output not listed

No provider description is available for this model yet.

snowflake/mistral-large 32K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-11-20 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-mini 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5-mini 128K context Input not listed Output not listed

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

google/gemma-3n-e4b-it 32.768K context $0.06/M input $0.12/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512 262.144K context $0.15/M input $0.15/M output

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

qwen/qwen3.6-35b-a3b 262.144K context $0.1/M input $0.9/M output

This model always redirects to the latest model in the GPT Luna family.

~openai/gpt-luna-latest 1.05M context $0.2/M input $1.2/M output

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

qwen/qwen3.5-9b:batch 262.144K context $0.17/M input $0.25/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2 1.04858M context $1.54/M input $4.84/M output

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

amazon/nova-pro-v1 300K context $0.8/M input $3.2/M output

No provider description is available for this model yet.

bedrock/ai21.j2-ultra-v1 8.191K context $18.8/M input $18.8/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.1 202.8K context $1.54/M input $4.84/M output