1,395 models

No provider description is available for this model yet.

novita/qwen/qwen3-vl-8b-instruct 131.072K context $0.08/M input $0.5/M output

No provider description is available for this model yet.

llamagate/gemma3-4b 128K context $0.03/M input $0.08/M output

No provider description is available for this model yet.

libertai/gemma-4-31b-it 262.144K context $0.15/M input $0.4/M output

No provider description is available for this model yet.

libertai/gemma-4-31b-it-thinking 262.144K context $0.15/M input $0.4/M output

No provider description is available for this model yet.

libertai/qwen3.6-27b 262.144K context $0.15/M input $0.5/M output

No provider description is available for this model yet.

libertai/qwen3.6-27b-thinking 262.144K context $0.15/M input $0.5/M output

No provider description is available for this model yet.

libertai/qwen3.6-35b-a3b 262.144K context $0.15/M input $0.5/M output

No provider description is available for this model yet.

libertai/qwen3.6-35b-a3b-thinking 262.144K context $0.15/M input $0.5/M output

No provider description is available for this model yet.

libertai/qwen3.5-122b-a10b 262.144K context $0.25/M input $1.75/M output

No provider description is available for this model yet.

libertai/qwen3.5-122b-a10b-thinking 262.144K context $0.25/M input $1.75/M output

No provider description is available for this model yet.

openai/gpt-5-search-api 272K context $1.25/M input $10/M output

No provider description is available for this model yet.

gemini/gemini-2.0-flash-lite-001 1.04858M context $0.075/M input $0.3/M output

No provider description is available for this model yet.

gemini/gemini-pro-latest 1.04858M context $1.25/M input $10/M output

No provider description is available for this model yet.

bedrock_mantle/google.gemma-4-31b 256K context $0.14/M input $0.4/M output

No provider description is available for this model yet.

bedrock_mantle/google.gemma-4-e2b 128K context $0.04/M input $0.08/M output

No provider description is available for this model yet.

bedrock_mantle/xai.grok-4.3 131.072K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

snowflake/claude-4-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-6 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-4-opus 200K context $5/M input $25/M output

No provider description is available for this model yet.

snowflake/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

snowflake/claude-3-7-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/openai-gpt-4.1 300K context $2/M input $8/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5 300K context $1.25/M input $10/M output

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo:batch 128K context $5/M input $15/M output

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3:batch 200K context $1/M input $4/M output

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

anthropic/claude-fable-5.1 1M context $10/M input $50/M output

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

openai/gpt-5.4-pro:batch 1.05M context $15/M input $90/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5.6-sol 1.05M context $2/M input $10/M output

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

google/gemini-3.8-flash:batch 1.04858M context $0.375/M input $1.875/M output

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

z-ai/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol:batch 1.05M context $1/M input $5/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output