3,330 models

No provider description is available for this model yet.

dashscope/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

dashscope/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

dashscope/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

qwen/qwen3.5-9b:batch 262.144K context $0.17/M input $0.25/M output

No provider description is available for this model yet.

dashscope/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

vertex_ai/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

gemini/gemini-3.7-flash 1.04858M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max 258.048K context Input not listed Output not listed

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

x-ai/grok-4.6 500K context $2/M input $6/M output

No provider description is available for this model yet.

anthropic/claude-3-opus-20240229 200K context $15/M input $75/M output

No provider description is available for this model yet.

groq/qwen/qwen3.6-27b 131.072K context $0.6/M input $3/M output

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

meta-llama/llama-3.3-70b-instruct:free 65.536K context Free input Free output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-preview 258.048K context Input not listed Output not listed

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

z-ai/glm-5.3 1.04858M context $1.4/M input $4.4/M output

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini:batch 200K context $0.55/M input $2.2/M output

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

mistralai/mistral-small-3.2-24b-instruct 256K context $0.094/M input $0.25/M output

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

anthropic/claude-haiku-4.5:batch 200K context $0.5/M input $2.5/M output

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite:batch 1.04858M context $0.15/M input $1.25/M output

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813:batch 1.04858M context $0.66/M input $1.98/M output