No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| z-ai/glm-4.5openrouter/z-ai/glm-4.5 | 131.072K | $0.6 | $2.2 | — | |||
| z-ai/glm-4.5-airopenrouter/z-ai/glm-4.5-air | 131.072K | $0.13 | $0.85 | — | |||
| moonshotai/kimi-k2openrouter/moonshotai/kimi-k2 | 131.072K | $0.57 | $2.3 | — | |||
| minimax/minimax-m1openrouter/minimax/minimax-m1 | 1M | $0.55 | $2.2 | — | |||
| openai/o3-proopenrouter/openai/o3-pro | 200K | $20 | $80 | — | |||
| google/gemini-2.5-pro-previewopenrouter/google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| mistralai/mistral-medium-3openrouter/mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| google/gemini-2.5-pro-preview-05-06openrouter/google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| meta-llama/llama-guard-4-12bopenrouter/meta-llama/llama-guard-4-12b | 163.84K | $0.18 | $0.18 | — | |||
| qwen/qwen3-30b-a3bopenrouter/qwen/qwen3-30b-a3b | 131.072K | $0.12 | $0.5 | — | |||
| qwen/qwen3-8bopenrouter/qwen/qwen3-8b | 131.072K | $0.117 | $0.455 | — | |||
| qwen/qwen3-14bopenrouter/qwen/qwen3-14b | 131.072K | $0.12 | $0.24 | — | |||
| qwen/qwen3-32bopenrouter/qwen/qwen3-32b | 131.072K | $0.08 | $0.28 | — | |||
| qwen/qwen3-235b-a22bopenrouter/qwen/qwen3-235b-a22b | 131.072K | $0.455 | $1.82 | — | |||
| openai/o4-mini-highopenrouter/openai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| meta-llama/llama-4-maverickopenrouter/meta-llama/llama-4-maverick | 1.04858M | $0.2 | $0.696 | — | |||
| meta-llama/llama-4-scoutopenrouter/meta-llama/llama-4-scout | 1.31072M | $0.1 | $0.3 | — | |||
| openai/o1-proopenrouter/openai/o1-pro | 200K | $150 | $600 | — | |||
| google/gemma-3-4b-itopenrouter/google/gemma-3-4b-it | 131.072K | $0.05 | $0.1 | — | |||
| google/gemma-3-12b-itopenrouter/google/gemma-3-12b-it | 131.072K | $0.05 | $0.15 | — | |||
| google/gemma-3-27b-itopenrouter/google/gemma-3-27b-it | 131.072K | $0.08 | $0.45 | — | |||
| mistralai/mistral-sabaopenrouter/mistralai/mistral-saba | 32.768K | $0.2 | $0.6 | — | |||
| qwen/qwen2.5-vl-72b-instructopenrouter/qwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — | |||
| qwen/qwen-plusopenrouter/qwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| mistralai/mistral-small-24b-instruct-2501openrouter/mistralai/mistral-small-24b-instruct-2501 | 32.768K | $0.05 | $0.08 | — | |||
| minimax/minimax-01openrouter/minimax/minimax-01 | 1.00019M | $0.2 | $1.1 | — | |||
| meta-llama/llama-3.3-70b-instructopenrouter/meta-llama/llama-3.3-70b-instruct | 131.072K | $0.1 | $0.32 | — | |||
| openai/gpt-4o-2024-11-20openrouter/openai/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| mistralai/mistral-large-2407openrouter/mistralai/mistral-large-2407 | 131.072K | $2 | $6 | — | |||
| qwen/qwen-2.5-7b-instructopenrouter/qwen/qwen-2.5-7b-instruct | 32.768K | $0.1 | $0.2 | — | |||
| meta-llama/llama-3.2-1b-instructopenrouter/meta-llama/llama-3.2-1b-instruct | 60K | $0.027 | $0.201 | — | |||
| meta-llama/llama-3.2-3b-instructopenrouter/meta-llama/llama-3.2-3b-instruct | 131.072K | $0.05 | $0.33 | — | |||
| qwen/qwen-2.5-72b-instructopenrouter/qwen/qwen-2.5-72b-instruct | 32.768K | $0.36 | $0.4 | — | |||
| openai/gpt-4o-2024-08-06openrouter/openai/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| meta-llama/llama-3.1-70b-instructopenrouter/meta-llama/llama-3.1-70b-instruct | 131.072K | $0.4 | $0.4 | — | |||
| meta-llama/llama-3.1-8b-instructopenrouter/meta-llama/llama-3.1-8b-instruct | 131.072K | $0.05 | $0.08 | — | |||
| mistralai/mistral-nemoopenrouter/mistralai/mistral-nemo | 131.072K | $0.019 | $0.03 | — | |||
| openai/gpt-4o-mini-2024-07-18openrouter/openai/gpt-4o-mini-2024-07-18 | 128K | $0.15 | $0.6 | — | |||
| openai/gpt-4-turboopenrouter/openai/gpt-4-turbo | 128K | $10 | $30 | — | |||
| openai/gpt-4-turbo-previewopenrouter/openai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch | 1.04858M | $0.25 | $1.5 | — | |||
| OpenAI: GPT-6 Astra (batch)openai/gpt-6-astra:batch | 1.05M | $5 | $25 | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| Qwen: Qwen3 VL 32B Instructqwen/qwen3-vl-32b-instruct | 131.072K | $0.104 | $0.416 | — | |||
| AionLabs: Aion-RP 1.0 (8B)aion-labs/aion-rp-llama-3.1-8b | 32.768K | $0.8 | $1.6 | — | |||
| Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1M | Free | Free | — | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — |