3,199 models

No provider description is available for this model yet.

mistral/zai-glm-5-2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openai/chat-latest 400K context $5/M input $30/M output

No provider description is available for this model yet.

github_copilot/gpt-5-mini 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4o-mini 64K context Input not listed Output not listed

No provider description is available for this model yet.

openai/daybreak-blue-latest 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-11-20 64K context Input not listed Output not listed

No provider description is available for this model yet.

snowflake/mistral-large 32K context Input not listed Output not listed

No provider description is available for this model yet.

openai/daybreak-red-latest 400K context $12.5/M input $75/M output

No provider description is available for this model yet.

snowflake/mistral-7b 32K context Input not listed Output not listed

No provider description is available for this model yet.

snowflake/llama3.3-70b 128K context $0.72/M input $0.72/M output

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1:batch 200K context $7.5/M input $30/M output

No provider description is available for this model yet.

snowflake/llama3.2-3b 128K context Input not listed Output not listed

LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI. It delivers strong chat quality in a 1.2B parameter footprint, with efficient edge inference and broad runtime support.

liquid/lfm-2.5-1.2b-instruct:free 32.768K context Free input Free output

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

liquid/lfm-2.5-2.6b:free 65.536K context Free input Free output

This model always redirects to the latest model in the GPT Astra family.

~openai/gpt-astra-latest 1.05M context $10/M input $50/M output

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

openai/gpt-oss-120b:free 131.072K context Free input Free output

This model always redirects to the latest model in the GPT Sol family.

~openai/gpt-sol-latest 1.05M context $2/M input $10/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-08-06 64K context Input not listed Output not listed

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2.5 260K context $0.04/M input $0.15/M output

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

mistralai/codestral-2508 256K context $0.3/M input $0.9/M output

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813:batch 1.04858M context $0.66/M input $1.98/M output

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini:batch 200K context $0.55/M input $2.2/M output

No provider description is available for this model yet.

github_copilot/gpt-4o-2024-05-13 64K context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/us-east-2/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output