No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast...
No provider description is available for this model yet.
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
No provider description is available for this model yet.
No provider description is available for this model yet.
Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
No provider description is available for this model yet.
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
No provider description is available for this model yet.
No provider description is available for this model yet.
KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us/gpt-5.6-solazure/us/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| us/gpt-5.6-terraazure/us/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| AionLabs: Aion-RP 1.0 (8B)aion-labs/aion-rp-llama-3.1-8b | 32.768K | $0.8 | $1.6 | — | |||
| mixtral-8x22B-Instruct-v0.1ollama/mixtral-8x22b-instruct-v0.1 | 65.536K | — | — | — | |||
| ap-south-1/moonshotai.kimi-k2.5bedrock/ap-south-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| us/gpt-5.6-lunaazure/us/gpt-5.6-luna | 1.05M | $1.1 | $6.6 | — | |||
| eu/gpt-5.6azure/eu/gpt-5.6 | 1.05M | $5.5 | $33 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| eu/gpt-5.6-solazure/eu/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.6-terraazure/eu/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.9 | — | |||
| us.anthropic.claude-3-5-haiku-20241022-v1:0bedrock/us.anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| us.anthropic.claude-opus-4-7bedrock_converse/us.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| gpt-5.5azure/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| global.anthropic.claude-opus-4-7bedrock_converse/global.anthropic.claude-opus-4-7 | 1M | $5 | $25 | — | |||
| us/gpt-5.5azure/us/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| eu/gpt-5.5azure/eu/gpt-5.5 | 1.05M | $5.5 | $33 | — | |||
| Mistral: Mistral Medium 3mistralai/mistral-medium-3 | 131.072K | $0.4 | $2 | — | |||
| chatgpt-4o-latestopenai/chatgpt-4o-latest | 128K | $5 | $15 | — | |||
| gpt-5.5-2026-04-23azure/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| llama-3.3-70bcerebras/llama-3.3-70b | 128K | $0.85 | $1.2 | — | |||
| us/gpt-5.5-2026-04-23azure/us/gpt-5.5-2026-04-23 | 1.05M | $5.5 | $33 | — | |||
| Tencent: Hy3 (free)tencent/hy3:free | 262.144K | Free | Free | — | |||
| inclusionAI: Ling-2.6-1Tinclusionai/ling-2.6-1t | 262.144K | $0.075 | $0.625 | — | |||
| openai/gpt-4-turboopenrouter/openai/gpt-4-turbo | 128K | $10 | $30 | — | |||
| OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batch | 1.05M | $0.1 | $0.6 | — | |||
| us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| ap-south-1/moonshotai.kimi-k2-thinkingbedrock/ap-south-1/moonshotai.kimi-k2-thinking | 262.144K | $0.71 | $2.94 | — | |||
| Arcee AI: Coder Largearcee-ai/coder-large | 32.768K | $0.5 | $0.8 | — | |||
| anthropic.claude-mythos-previewbedrock/anthropic.claude-mythos-preview | 1M | — | — | — | |||
| ap-south-1/minimax.minimax-m2.5bedrock/ap-south-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| SpaceXAI: Grok Build 0.1x-ai/grok-build-0.1 | 256K | $1 | $2 | — | |||
| anthropic.claude-opus-4-7bedrock_converse/anthropic.claude-opus-4-7 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| Kwaipilot: KAT-Coder-Air V2.5 (free)kwaipilot/kat-coder-air-v2.5:free | 256K | Free | Free | — | |||
| MiniMaxAI/MiniMax-M2.1gmi/minimaxai/minimax-m2.1 | 196.608K | $0.3 | $1.2 | — | |||
| moonshotai/Kimi-K2-Thinkinggmi/moonshotai/kimi-k2-thinking | 262.144K | $0.8 | $1.2 | — | |||
| Kwaipilot: KAT-Coder-Pro V2.5 (free)kwaipilot/kat-coder-pro-v2.5:free | 256K | Free | Free | — | |||
| Qwen: Qwen3.7 Maxqwen/qwen3.7-max | 1M | $1.475 | $4.425 | — | |||
| openai/gpt-4o-mini-2024-07-18openrouter/openai/gpt-4o-mini-2024-07-18 | 128K | $0.15 | $0.6 | — | |||
| openai/o3-proopenrouter/openai/o3-pro | 200K | $20 | $80 | — | |||
| us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/us-gov-east-1/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| au.anthropic.claude-opus-4-6-v1bedrock_converse/au.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| google/gemini-3-flash-previewgmi/google/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | — | |||
| OpenAI: GPT-5.5 Pro (batch)openai/gpt-5.5-pro:batch | 1.05M | $15 | $90 | — | |||
| Anthropic: Claude Opus 4.7 (Fast)anthropic/claude-opus-4.7-fast | 1M | $30 | $150 | — | |||
| google/gemini-3-pro-previewgmi/google/gemini-3-pro-preview | 1.04858M | $2 | $12 | — | |||
| deepseek-ai/DeepSeek-V3-0324gmi/deepseek-ai/deepseek-v3-0324 | 163.84K | $0.28 | $0.88 | — | |||
| Nous: Hermes 3 405B Instruct (free)nousresearch/hermes-3-llama-3.1-405b:free | 131.072K | Free | Free | — | |||
| TheDrummer: Cydonia 24B V4.1thedrummer/cydonia-24b-v4.1 | 131.072K | $0.3 | $0.5 | — |