4,248 models

This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.

openai/gpt-3.5-turbo-instruct 4.095K context $1.5/M input $2/M output

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...

openai/gpt-3.5-turbo-16k 16.385K context $3/M input $4/M output

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

mistralai/mistral-large-2407 131.072K context $2/M input $6/M output

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...

qwen/qwen-2.5-coder-32b-instruct 32.768K context $0.66/M input $1/M output

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

thedrummer/unslopnemo-12b 1.024M context $0.4/M input $0.4/M output

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

anthracite-org/magnum-v4-72b 32.768K context $2.5/M input $5/M output

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

qwen/qwen-2.5-7b-instruct 32.768K context $0.1/M input $0.2/M output

An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory. Meant for use in roleplay/narrative situations.

mancer/weaver 8K context $0.4/M input $0.75/M output

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...

mistralai/mixtral-8x22b-instruct 65.536K context $2/M input $6/M output

WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...

microsoft/wizardlm-2-8x22b 65.535K context $0.62/M input $0.62/M output

GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.

openai/gpt-3.5-turbo-0613 4.095K context $1/M input $2/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.1 204.8K context $1.38/M input $4.4/M output

No provider description is available for this model yet.

openrouter/mancer/weaver 8K context $5.625/M input $5.625/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.6 262.144K context $0.8/M input $3.4/M output

No provider description is available for this model yet.

openrouter/gryphe/mythomax-l2-13b 8.192K context $1.875/M input $1.875/M output

No provider description is available for this model yet.

novita/qwen/qwen3.6-27b 262.144K context $0.6/M input $3.6/M output

No provider description is available for this model yet.

novita/xiaomimimo/mimo-v2.5-pro 1.04858M context $0.522/M input $1.044/M output

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

inclusionai/ling-3.0-flash-vl:free 262.144K context Free input Free output

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

kwaipilot/kat-coder-pro-v2.5:free 256K context Free input Free output

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

kwaipilot/kat-coder-air-v2.5:free 256K context Free input Free output

No provider description is available for this model yet.

novita/qwen/qwen3.7-max 1M context $1.25/M input $3.75/M output

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra:batch 1.05M context $5/M input $25/M output

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

tencent/hy3:free 262.144K context Free input Free output

Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling and reasoning, with a 256K...

poolside/laguna-m.1:free 262.144K context Free input Free output

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

x-ai/grok-4.5 500K context $2/M input $6/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3 524.288K context $0.3/M input $1.2/M output