3,229 models

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash:batch 1.04858M context $0.75/M input $4.5/M output

No provider description is available for this model yet.

openrouter/openai/gpt-5-pro 400K context $15/M input $120/M output

No provider description is available for this model yet.

bedrock/amazon.titan-text-lite-v1 42K context $0.3/M input $0.4/M output

No provider description is available for this model yet.

gmi/deepseek-ai/deepseek-v3.2 163.84K context $0.28/M input $0.4/M output

No provider description is available for this model yet.

databricks/databricks-inkling 1M context $1/M input $4.05/M output

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

google/gemini-3.1-flash-lite:batch 1.04858M context $0.125/M input $0.75/M output

No provider description is available for this model yet.

databricks/databricks-grok-4-6 500K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

qwen_ai_platform/qwq-plus 98.304K context $0.8/M input $2.4/M output

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

openai/o1-pro:batch 200K context $75/M input $300/M output

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

inclusionai/ling-3.0-flash-vl 131.072K context $0.06/M input $0.18/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512 262.144K context $0.15/M input $0.15/M output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1:batch 200K context $7.5/M input $37.5/M output

Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling and reasoning, with a 256K...

poolside/laguna-m.1:free 262.144K context Free input Free output

No provider description is available for this model yet.

cerebras/zai-glm-4.7 128K context $2.25/M input $2.75/M output

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

meta-llama/llama-3.2-11b-vision-instruct 131.072K context $0.345/M input $0.345/M output

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.

thedrummer/cydonia-24b-v4.1 131.072K context $0.3/M input $0.5/M output

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

nousresearch/hermes-3-llama-3.1-405b:free 131.072K context Free input Free output

No provider description is available for this model yet.

gmi/deepseek-ai/deepseek-v3-0324 163.84K context $0.28/M input $0.88/M output

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.5-pro:batch 1.05M context $15/M input $90/M output

No provider description is available for this model yet.

gmi/google/gemini-3-pro-preview 1.04858M context $2/M input $12/M output

No provider description is available for this model yet.

gmi/google/gemini-3-flash-preview 1.04858M context $0.5/M input $3/M output