Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-72b-instruct 32.768K context $0.36/M input $0.4/M output

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7 1M context $5/M input $25/M output

No provider description is available for this model yet.

xai/grok-4.20-multi-agent-0309 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

anthropic/claude-mythos-preview 1M context $10/M input $50/M output

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...

mistralai/mixtral-8x22b-instruct 65.536K context $2/M input $6/M output

No provider description is available for this model yet.

mistral/labs-leanstral-1-5 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o 64K context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/us-east-2/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-05-13 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-08-06 64K context Input not listed Output not listed

This model always redirects to the latest model in the GPT Sol family.

~openai/gpt-sol-latest 1.05M context $2/M input $10/M output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

This model always redirects to the latest model in the GPT Astra family.

~openai/gpt-astra-latest 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

snowflake/llama3.2-3b 128K context Input not listed Output not listed

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1:batch 200K context $7.5/M input $30/M output

No provider description is available for this model yet.

snowflake/llama3.3-70b 128K context $0.72/M input $0.72/M output

No provider description is available for this model yet.

snowflake/mistral-7b 32K context Input not listed Output not listed

No provider description is available for this model yet.

snowflake/mistral-large 32K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-11-20 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-mini 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5-mini 128K context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/fw-nemotron-3-ultra-nvfp4 262.144K context $0.6/M input $2.4/M output

This model always redirects to the latest model in the GPT Luna family.

~openai/gpt-luna-latest 1.05M context $0.2/M input $1.2/M output

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-7b-instruct 32.768K context $0.1/M input $0.2/M output

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

qwen/qwen3.5-9b:batch 262.144K context $0.17/M input $0.25/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview 1.04858M context $1.25/M input $10/M output

No provider description is available for this model yet.

azure_ai/fw-minimax-m3 512K context $0.33/M input $1.32/M output