735 models

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

qwen/qwen3.6-flash 1M context $0.188/M input $1.125/M output

No provider description is available for this model yet.

xai/grok-4-fast-reasoning 2M context $0.2/M input $0.5/M output

No provider description is available for this model yet.

xai/grok-4-fast-non-reasoning 2M context $0.2/M input $0.5/M output

No provider description is available for this model yet.

xai/grok-4-1-fast 2M context $0.2/M input $0.5/M output

No provider description is available for this model yet.

xai/grok-4-1-fast-reasoning 2M context $0.2/M input $0.5/M output

No provider description is available for this model yet.

xai/grok-4-1-fast-non-reasoning 2M context $0.2/M input $0.5/M output

No provider description is available for this model yet.

xai/grok-4.3-latest 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

novita/minimaxai/minimax-m1-80k 1M context $0.55/M input $2.2/M output

No provider description is available for this model yet.

gemini/gemini-2.0-flash-lite-001 1.04858M context $0.075/M input $0.3/M output

No provider description is available for this model yet.

gemini/gemini-pro-latest 1.04858M context $1.25/M input $10/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-mini 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-nano 5M context $0.15/M input $0.6/M output

No provider description is available for this model yet.

tencent/deepseek-v4-pro 1M context $0.435/M input $0.87/M output

No provider description is available for this model yet.

tencent/deepseek-v4-flash 1M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

pinstripes/ps/minimax-m2.7 1.00019M context $0.255/M input $0.55/M output

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

google/gemini-3.7-flash:batch 1.04858M context $0.375/M input $1.875/M output

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

nvidia/nemotron-3-ultra-550b-a55b:free 1M context Free input Free output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5:batch 1.05M context $2.5/M input $15/M output

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

nvidia/nemotron-3.5-lightning:free 1M context Free input Free output

Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

sakana/fugu-ultra-v2 1M context $5/M input $30/M output

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra:batch 1.05M context $5/M input $25/M output