This model is a variant of GPT-3.5 Turbo tuned for instructional prompts and omitting chat-related optimizations. Training data: up to Sep 2021.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up...
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...
An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory. Meant for use in roleplay/narrative situations.
Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...
WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
No provider description is available for this model yet.
No provider description is available for this model yet.
KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
No provider description is available for this model yet.
KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
No provider description is available for this model yet.
No provider description is available for this model yet.
Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai/), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling and reasoning, with a 256K...
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| OpenAI: GPT-3.5 Turbo Instructopenai/gpt-3.5-turbo-instruct | 4.095K | $1.5 | $2 | — | |||
| OpenAI: GPT-3.5 Turbo 16kopenai/gpt-3.5-turbo-16k | 16.385K | $3 | $4 | — | |||
| Mistral Large 2407mistralai/mistral-large-2407 | 131.072K | $2 | $6 | — | |||
| Qwen2.5 Coder 32B Instructqwen/qwen-2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| TheDrummer: UnslopNemo 12Bthedrummer/unslopnemo-12b | 1.024M | $0.4 | $0.4 | — | |||
| Magnum v4 72Banthracite-org/magnum-v4-72b | 32.768K | $2.5 | $5 | — | |||
| Qwen: Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct | 32.768K | $0.1 | $0.2 | — | |||
| Mancer: Weaver (alpha)mancer/weaver | 8K | $0.4 | $0.75 | — | |||
| Mistral: Mixtral 8x22B Instructmistralai/mixtral-8x22b-instruct | 65.536K | $2 | $6 | — | |||
| WizardLM-2 8x22Bmicrosoft/wizardlm-2-8x22b | 65.535K | $0.62 | $0.62 | — | |||
| OpenAI: GPT-3.5 Turbo (older v0613)openai/gpt-3.5-turbo-0613 | 4.095K | $1 | $2 | — | |||
| zai-org/glm-5.1novita/zai-org/glm-5.1 | 204.8K | $1.38 | $4.4 | — | |||
| mancer/weaveropenrouter/mancer/weaver | 8K | $5.625 | $5.625 | — | |||
| accounts/fireworks/models/deepseek-v3p1fireworks_ai/accounts/fireworks/models/deepseek-v3p1 | 128K | $0.56 | $1.68 | — | |||
| Qwen/Qwen2.5-VL-72B-Instructtogether_ai/qwen/qwen2.5-vl-72b-instruct | Not documented | $1.95 | $8 | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — | |||
| gryphe/mythomax-l2-13bopenrouter/gryphe/mythomax-l2-13b | 8.192K | $1.875 | $1.875 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-26000microsoft/Dayhoff-170M-GRS-SS-26000 | Not documented | — | — | — | |||
| qwen/qwen3.6-27bnovita/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| google/gemini-3.1-pro-previewopenrouter/google/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | — | |||
| accounts/fireworks/models/deepseek-v3-0324fireworks_ai/accounts/fireworks/models/deepseek-v3-0324 | 163.84K | $0.9 | $0.9 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-50000microsoft/Dayhoff-170M-GRS-SS-50000 | Not documented | — | — | — | |||
| xiaomimimo/mimo-v2.5-pronovita/xiaomimimo/mimo-v2.5-pro | 1.04858M | $0.522 | $1.044 | — | |||
| google/gemini-3.1-flash-liteopenrouter/google/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-62000microsoft/Dayhoff-170M-GRS-SS-62000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-SS-74000microsoft/Dayhoff-170M-GRS-SS-74000 | Not documented | — | — | — | |||
| mistralai/ministral-14b-2512openrouter/mistralai/ministral-14b-2512 | 262.144K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/deepseek-v3fireworks_ai/accounts/fireworks/models/deepseek-v3 | 128K | $0.9 | $0.9 | — | |||
| inclusionAI: Ling 3.0 Flash VL (free)inclusionai/ling-3.0-flash-vl:free | 262.144K | Free | Free | — | |||
| microsoft/Dayhoff-170M-GRS-26000microsoft/Dayhoff-170M-GRS-26000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GRS-SS-98000microsoft/Dayhoff-170M-GRS-SS-98000 | Not documented | — | — | — | |||
| Kwaipilot: KAT-Coder-Pro V2.5 (free)kwaipilot/kat-coder-pro-v2.5:free | 256K | Free | Free | — | |||
| nvidia/Riva-Translate-4B-Instruct-v1.1nvidia/Riva-Translate-4B-Instruct-v1.1 | Not documented | — | — | — | |||
| Kwaipilot: KAT-Coder-Air V2.5 (free)kwaipilot/kat-coder-air-v2.5:free | 256K | Free | Free | — | |||
| nvidia/Nemotron-3-Labs-Ultra-Math-SFTnvidia/Nemotron-3-Labs-Ultra-Math-SFT | Not documented | — | — | — | |||
| ap-south-1/minimax.minimax-m2.5bedrock/ap-south-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| ap-south-1/moonshotai.kimi-k2-thinkingbedrock/ap-south-1/moonshotai.kimi-k2-thinking | 262.144K | $0.71 | $2.94 | — | |||
| nvidia/Nemotron-3-Labs-Ultra-Math-RLnvidia/Nemotron-3-Labs-Ultra-Math-RL | Not documented | — | — | — | |||
| qwen/qwen3.7-maxnovita/qwen/qwen3.7-max | 1M | $1.25 | $3.75 | — | |||
| OpenAI: GPT-6 Astra (batch)openai/gpt-6-astra:batch | 1.05M | $5 | $25 | — | |||
| Tencent: Hy3 (free)tencent/hy3:free | 262.144K | Free | Free | — | |||
| google/gemini-3.1-flash-lite-previewopenrouter/google/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 | $1.5 | — | |||
| ap-south-1/moonshotai.kimi-k2.5bedrock/ap-south-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| Poolside: Laguna M.1 (free)poolside/laguna-m.1:free | 262.144K | Free | Free | — | |||
| ap-south-1/qwen.qwen3-coder-nextbedrock/ap-south-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| ap-southeast-2/minimax.minimax-m2.5bedrock/ap-southeast-2/minimax.minimax-m2.5 | 1M | $0.309 | $1.236 | — | |||
| SpaceXAI: Grok 4.5x-ai/grok-4.5 | 500K | $2 | $6 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-110000microsoft/Dayhoff-170M-GRS-SS-110000 | Not documented | — | — | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-122000microsoft/Dayhoff-170M-GRS-SS-122000 | Not documented | — | — | — |