3,310 models

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

anthropic/claude-sonnet-4 200K context $3/M input $15/M output

No provider description is available for this model yet.

oci/openai.gpt-5 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

qwencloud/qwen3-vl-32b-instruct 131.072K context $0.16/M input $0.64/M output

No provider description is available for this model yet.

qwencloud/qwen3-vl-32b-thinking 131.072K context $0.16/M input $2.87/M output

No provider description is available for this model yet.

qwencloud/qwen3-vl-plus 260.096K context Input not listed Output not listed

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash:batch 1.04858M context $0.75/M input $4.5/M output

No provider description is available for this model yet.

qwencloud/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

qwencloud/qwen3.7-max 991.808K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwencloud/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

oci/xai.grok-code-fast-1 131.072K context $5/M input $25/M output

No provider description is available for this model yet.

dashscope/qwen-turbo-latest 1M context $0.05/M input $0.2/M output

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

qwen/qwen3.8-flash 1M context $0.15/M input $0.47/M output

No provider description is available for this model yet.

qwencloud/qwen3.8-max 991.808K context $2/M input $6/M output

No provider description is available for this model yet.

qwencloud/qwq-plus 98.304K context $0.8/M input $2.4/M output

No provider description is available for this model yet.

azure/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

qwen/qwen3-14b 40.96K context $0.12/M input $0.24/M output

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

openai/o4-mini:batch 200K context $0.55/M input $2.2/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.1 202.745K context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

qwen_ai_platform/kimi-k2.7-code 229.376K context $0.95/M input $4/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-coder 1M context $0.3/M input $1.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

deepinfra/google/gemma-4-31b-it 262.144K context $0.13/M input $0.38/M output

No provider description is available for this model yet.

oci/xai.grok-4.20-multi-agent 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus 129.024K context $0.4/M input $1.2/M output

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

sakana/fugu-ultra-v2 1M context $5/M input $30/M output

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano:batch 1.04758M context $0.05/M input $0.2/M output

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...

relace/relace-apply-3 256K context $0.85/M input $1.25/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-07-28 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-2025-09-11 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-plus-latest 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen-turbo 129.024K context $0.05/M input $0.2/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-4.7 202.752K context $0.4/M input $1.75/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-30b-a3b 129.024K context Input not listed Output not listed

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

minimax/minimax-m2.7 204.8K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-flash 997.952K context Input not listed Output not listed

No provider description is available for this model yet.

qwen_ai_platform/qwen3-coder-plus 997.952K context Input not listed Output not listed