3,327 models

No provider description is available for this model yet.

openrouter/z-ai/glm-4.5 131.072K context $0.6/M input $2.2/M output

No provider description is available for this model yet.

openrouter/z-ai/glm-4.5-air 131.072K context $0.13/M input $0.85/M output

No provider description is available for this model yet.

openrouter/moonshotai/kimi-k2 131.072K context $0.57/M input $2.3/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-m1 1M context $0.55/M input $2.2/M output

No provider description is available for this model yet.

openrouter/openai/o3-pro 200K context $20/M input $80/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-30b-a3b 131.072K context $0.12/M input $0.5/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-8b 131.072K context $0.117/M input $0.455/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-14b 131.072K context $0.12/M input $0.24/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-32b 131.072K context $0.08/M input $0.28/M output

No provider description is available for this model yet.

openrouter/qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

No provider description is available for this model yet.

openrouter/openai/o4-mini-high 200K context $1.1/M input $4.4/M output

No provider description is available for this model yet.

openrouter/openai/o1-pro 200K context $150/M input $600/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-4b-it 131.072K context $0.05/M input $0.1/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-12b-it 131.072K context $0.05/M input $0.15/M output

No provider description is available for this model yet.

openrouter/google/gemma-3-27b-it 131.072K context $0.08/M input $0.45/M output

No provider description is available for this model yet.

openrouter/mistralai/mistral-saba 32.768K context $0.2/M input $0.6/M output

No provider description is available for this model yet.

openrouter/qwen/qwen-plus 1M context $0.26/M input $0.78/M output

No provider description is available for this model yet.

openrouter/minimax/minimax-01 1.00019M context $0.2/M input $1.1/M output

No provider description is available for this model yet.

openrouter/mistralai/mistral-nemo 131.072K context $0.019/M input $0.03/M output

No provider description is available for this model yet.

openrouter/openai/gpt-4-turbo 128K context $10/M input $30/M output

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

google/gemini-3-flash-preview:batch 1.04858M context $0.25/M input $1.5/M output

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra:batch 1.05M context $5/M input $25/M output

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

inclusionai/ling-3.0-flash-fin:free 262.144K context Free input Free output

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

qwen/qwen3-vl-32b-instruct 131.072K context $0.104/M input $0.416/M output

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

aion-labs/aion-rp-llama-3.1-8b 32.768K context $0.8/M input $1.6/M output

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

qwen/qwen2.5-vl-72b-instruct 128K context $0.8/M input $1/M output

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini:batch 128K context $0.075/M input $0.3/M output

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

nvidia/nemotron-3-ultra-550b-a55b:free 1M context Free input Free output

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash:batch 1.04858M context $0.375/M input $1.875/M output

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

anthropic/claude-sonnet-4.6 1M context $3/M input $15/M output