Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

anthropic/claude-opus-4.7 1M context $5/M input $25/M output

Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B. It introduces the same hybrid mode as the larger 405B release, allowing the model to either...

nousresearch/hermes-4-70b 131.072K context $0.13/M input $0.4/M output

No provider description is available for this model yet.

bedrock/us-east-1/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

anthropic/claude-mythos-preview 1M context $10/M input $50/M output

No provider description is available for this model yet.

mistral/labs-leanstral-1-5 262.144K context Input not listed Output not listed

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

openai/o4-mini:batch 200K context $0.55/M input $2.2/M output

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

qwen/qwen3-next-80b-a3b-thinking 262.144K context $0.15/M input $1.2/M output

No provider description is available for this model yet.

github_copilot/gpt-4o 64K context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/us-east-2/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-05-13 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-08-06 64K context Input not listed Output not listed

This model always redirects to the latest model in the GPT Astra family.

~openai/gpt-astra-latest 1.05M context $10/M input $50/M output

No provider description is available for this model yet.

snowflake/llama3.2-3b 128K context Input not listed Output not listed

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1:batch 200K context $7.5/M input $30/M output

No provider description is available for this model yet.

snowflake/llama3.3-70b 128K context $0.72/M input $0.72/M output

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

nvidia/nemotron-nano-9b-v2:free 128K context Free input Free output

No provider description is available for this model yet.

snowflake/mistral-large 32K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-11-20 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-mini 64K context Input not listed Output not listed

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini:batch 200K context $0.55/M input $2.2/M output

No provider description is available for this model yet.

github_copilot/gpt-5 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5-mini 128K context Input not listed Output not listed

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash:batch 1.04858M context $0.075/M input $0.25/M output

This model always redirects to the latest model in the GPT Luna family.

~openai/gpt-luna-latest 1.05M context $0.2/M input $1.2/M output

No provider description is available for this model yet.

bedrock/ai21.jamba-1-5-mini-v1:0 256K context $0.2/M input $0.4/M output

No provider description is available for this model yet.

bedrock/ai21.jamba-instruct-v1:0 70K context $0.5/M input $0.7/M output

This model always redirects to the latest model in the GLM Flash family.

~z-ai/glm-flash-latest 1.04858M context $0.075/M input $0.25/M output

No provider description is available for this model yet.

openrouter/poolside/laguna-xs-2.1 262.144K context $0.06/M input $0.12/M output

No provider description is available for this model yet.

nebius/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output