2,861 models

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

openai/o3-mini-high:batch 200K context $0.55/M input $2.2/M output

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

google/gemini-3.1-flash-lite:batch 1.04858M context $0.125/M input $0.75/M output

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

qwen/qwen3-next-80b-a3b-instruct:free 262.144K context Free input Free output

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

inclusionai/ling-3.0-flash 262.144K context $0.021/M input $0.063/M output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:batch 262.144K context $0.39/M input $0.97/M output

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

x-ai/grok-4.20-multi-agent 2M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

bedrock/sa-east-1/deepseek.v3.2 163.84K context $0.74/M input $2.22/M output

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...

openai/gpt-5.2-pro:batch 400K context $10.5/M input $84/M output

No provider description is available for this model yet.

friendliai/google/gemma-4-31b-it 262.144K context $0.14/M input $0.4/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...

stealth/ox-alpha 1.04858M context Input not listed Output not listed

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

qwen/qwen-plus-2025-07-28:thinking 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

friendliai/minimaxai/minimax-m2.5 196.608K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

azure_ai/gpt-6-astra 922K context $10/M input $50/M output

No provider description is available for this model yet.

openai/gpt-5.6-cyber 400K context $12.5/M input $75/M output

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

openai/gpt-oss-120b:free 131.072K context Free input Free output

No provider description is available for this model yet.

openai/daybreak-red-latest 400K context $12.5/M input $75/M output

No provider description is available for this model yet.

openai/daybreak-blue-latest 1.05M context $5/M input $30/M output

No provider description is available for this model yet.

openai/chat-latest 400K context $5/M input $30/M output

No provider description is available for this model yet.

mistral/zai-glm-5-2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

mistral/glm-5-2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

openrouter/deepseek/deepseek-v4-pro 1.04858M context $1.32/M input $3.96/M output

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.6 500K context $2.2/M input $6.6/M output

No provider description is available for this model yet.

bedrock_converse/us.xai.grok-4.6 500K context $2.2/M input $6.6/M output

No provider description is available for this model yet.

azure_ai/fw-deepseek-v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

azure_ai/fw-deepseek-v4-pro 1M context $1.925/M input $3.828/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5 200K context $1.1/M input $3.52/M output

No provider description is available for this model yet.

bedrock/us-east-1/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.1 202.8K context $1.54/M input $4.84/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2 1.04858M context $1.54/M input $4.84/M output