1,808 models

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

nvidia/nemotron-3.5-lightning:free 1M context Free input Free output

No provider description is available for this model yet.

gemini/gemini-2.0-flash-lite-001 1.04858M context $0.075/M input $0.3/M output

No provider description is available for this model yet.

novita/zai-org/glm-4.7 204.8K context $0.6/M input $2.2/M output

No provider description is available for this model yet.

databricks/databricks-gpt-5-4 272K context $2.5/M input $15/M output

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

No provider description is available for this model yet.

databricks/databricks-gpt-5-2 272K context $1.75/M input $14/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-preview 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-plus 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-max-2026-01-23 258.048K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-vl-plus 260.096K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

databricks/databricks-gemini-3-pro 1.04858M context $2.5/M input $15/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.3 1.04858M context $1.26/M input $3.96/M output

No provider description is available for this model yet.

zai/glm-5.2 1M context $1.4/M input $4.4/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

openai/gpt-5.6-terra:batch 1.05M context $1/M input $6/M output

No provider description is available for this model yet.

openai/gpt-5-search-api 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

novita/minimax/minimax-m2.1 204.8K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-35b-a3b 262.144K context $0.248/M input $1.485/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-latest 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed